MCP & Gọi công cụ
mtrnix/metronix-memory
Self-hosted memory infrastructure for AI agents featuring MCP-native integration, hybrid RAG, a temporal knowledge graph, and an ontology layer, designed for local-model friendliness and durable, agent-scoped context.
★ 39⑂ 7Python
Apache-2.0Q86
MCP & Gọi công cụ
mycelium-hq/ai-brain-starter
An operating system and verification harness for Claude Code that provides persistent memory, accountability, journaling, and knowledge graph capabilities, ensuring context compounds across sessions rather than corrupting.
★ 33⑂ 25Python
MITQ86
Tác nhân & Đa tác nhân
pwrdrvr/pwragent
An open-source desktop coding agent that runs locally on a laptop and can be driven remotely via messaging platforms like Telegram, Discord, Slack, Mattermost, Feishu, Lark, or LINE.
★ 29⑂ 4TypeScript
MITQ86
MCP & Gọi công cụ
gethuman-sh/human
An open-source AI Software Factory that integrates tickets, docs, designs, and analytics into a secure pipeline, outputting shipped code via Claude-driven agents.
★ 62⑂ 6Go
MITQ85
Đánh giá, khả năng quan sát & an toàn
NVIDIA/garak
the LLM vulnerability scanner
★ 8,8K⑂ 1,2KPython
Apache-2.0Q84
Tác nhân & Đa tác nhân
anthropics/claude-agent-sdk-python
An open-source project focused on python sdk for agent applications.
★ 7,9K⑂ 1,2KPython
MITQ84
LLM & Mô hình nền tảng
QwenLM/Qwen3
Qwen3 is the large language model series developed by Qwen team, Alibaba Cloud
★ 27,5K⑂ 2KPython
Không phát hiện giấy phépQ82
Đánh giá, khả năng quan sát & an toàn
openai/evals
Evals is a framework for evaluating LLMs and LLM systems, and an open-source registry of benchmarks
★ 19,2K⑂ 3,1KPython
NOASSERTIONQ82
LLM & Mô hình nền tảng
QwenLM/Qwen3-Coder
Qwen3-Coder is the code version of Qwen3, the large language model series developed by Qwen team
★ 16,8K⑂ 1,2KPython
Không phát hiện giấy phépQ82
Suy luận, Triển khai & Thời gian chạy
NVIDIA/TensorRT-LLM
TensorRT LLM provides users with an easy-to-use Python API to define Large Language Models (LLMs) and supports state-of-the-art optimizations to perform inference efficiently on NVIDIA GPUs. TensorRT LLM also contains components to create Python and C++ runtimes that orchestrate the inference execution in a performant way
★ 14,4K⑂ 2,7KC++
NOASSERTIONQ82