Agenten & Multi-Agenten
eosphoros-ai/db-gpt
DB-GPT is a Python-based platform for building and operating AI data assistants that connect to structured and unstructured data, generate SQL and Python code, execute analysis workflows, and produce reports, charts, dashboards, and summaries.
★ 19,6K⑂ 2,9KPython
MITQ98
Agenten & Multi-Agenten
alibaba/open-code-review
An open-source AI code review CLI from Alibaba that combines deterministic review pipelines with an LLM agent to produce structured, line-level feedback.
★ 18K⑂ 1,2KGo
Apache-2.0Q98
MCP & Tool-Aufrufe
composiohq/composio
Composio is an open-source SDK monorepo for connecting AI agents to more than 1,000 toolkits. It provides tool discovery and execution, per-user sessions, authentication, triggers, context management, hosted MCP endpoints, provider adapters, a CLI, and a sandboxed workbench.
★ 29,5K⑂ 4,7KTypeScript
MITQ98
Agenten & Multi-Agenten
mlflow/mlflow
An open-source AI engineering platform for tracing, evaluating, monitoring, optimizing, governing, and deploying agents, LLM applications, and machine-learning models.
★ 27,3K⑂ 6,1KPython
Apache-2.0Q98
Agenten & Multi-Agenten
alibaba/zvec
Zvec is an Apache-2.0-licensed vector database designed to run directly inside applications. It supports dense and sparse vectors, full-text and hybrid search, structured filtering, local durable storage, and SDKs for several programming languages.
★ 15,4K⑂ 973C++
Apache-2.0Q98
Agenten & Multi-Agenten
neuml/txtai
txtai is a Python-based AI framework for semantic and vector search, retrieval-augmented generation, LLM orchestration, autonomous agents and multimodal language-model workflows.
★ 12,8K⑂ 851Python
Apache-2.0Q98
Inferenz, Bereitstellung & Laufzeit
kvcache-ai/mooncake
Mooncake is a C++ infrastructure project for large-scale LLM inference and training. It separates prefill, decode, and storage resources while providing high-performance transfer, distributed KV-cache storage, and fault-tolerant expert-parallel communication.
★ 6,1K⑂ 1KC++
Apache-2.0Q98
Inferenz, Bereitstellung & Laufzeit
gpustack/gpustack
GPUStack is an open-source GPU cluster manager for serving AI models and provisioning on-demand, SSH-accessible GPU instances. It orchestrates inference engines including vLLM, SGLang, and TensorRT-LLM across on-premises, Kubernetes, and cloud environments.
★ 5,4K⑂ 604Python
Apache-2.0Q98
Agenten & Multi-Agenten
giskard-ai/giskard-oss
An open-source Python library for evaluating and testing LLM-based and agentic systems, including multi-turn evaluations, LLM-as-judge checks and automated vulnerability scanning.
★ 5,7K⑂ 513Python
Apache-2.0Q98
Inferenz, Bereitstellung & Laufzeit
llm-d/llm-d
llm-d is an open-source serving stack that adds distributed orchestration, routing, cache management, autoscaling, and batch processing around model servers such as vLLM and SGLang. It targets high-scale production inference on Kubernetes and modern hardware accelerators.
★ 4K⑂ 649Shell
Apache-2.0Q98
Agenten & Multi-Agenten
langwatch/langwatch
An open-core platform for evaluating, testing, tracing, and monitoring LLM applications and AI agents before release and in production.
★ 3,5K⑂ 345TypeScript
Apache-2.0Q98
Inferenz, Bereitstellung & Laufzeit
vllm-project/vllm-ascend
A community-maintained hardware plugin that enables vLLM to run model inference and serving workloads on supported Ascend NPUs.
★ 2,7K⑂ 2,1KC++
Apache-2.0Q98