ToolAI.io · GitHub Channel

New AI Projects on GitHub

Discover newly added open-source AI projects on GitHub with repository facts, setup notes, licenses and related resources.

Public repository data

Project index

130 projects
Agents & Multi-Agent

DB-GPT: Open-Source Agentic AI Data Assistant

eosphoros-ai/db-gpt

DB-GPT is a Python-based platform for building and operating AI data assistants that connect to structured and unstructured data, generate SQL and Python code, execute analysis workflows, and produce reports, charts, dashboards, and summaries.

★ 19.6K⑂ 2.9KPython
MITQ98
Agents & Multi-Agent

OpenCodeReview

alibaba/open-code-review

An open-source AI code review CLI from Alibaba that combines deterministic review pipelines with an LLM agent to produce structured, line-level feedback.

★ 18K⑂ 1.2KGo
Apache-2.0Q98
MCP & Tool Calling

Composio

composiohq/composio

Composio is an open-source SDK monorepo for connecting AI agents to more than 1,000 toolkits. It provides tool discovery and execution, per-user sessions, authentication, triggers, context management, hosted MCP endpoints, provider adapters, a CLI, and a sandboxed workbench.

★ 29.5K⑂ 4.7KTypeScript
MITQ98
Agents & Multi-Agent

MLflow

mlflow/mlflow

An open-source AI engineering platform for tracing, evaluating, monitoring, optimizing, governing, and deploying agents, LLM applications, and machine-learning models.

★ 27.3K⑂ 6.1KPython
Apache-2.0Q98
Agents & Multi-Agent

Zvec: An Embedded, In-Process Vector Database

alibaba/zvec

Zvec is an Apache-2.0-licensed vector database designed to run directly inside applications. It supports dense and sparse vectors, full-text and hybrid search, structured filtering, local durable storage, and SDKs for several programming languages.

★ 15.4K⑂ 973C++
Apache-2.0Q98
Inference, Deployment & Runtime

Mooncake: KVCache-Centric Infrastructure for Distributed LLM Serving

kvcache-ai/mooncake

Mooncake is a C++ infrastructure project for large-scale LLM inference and training. It separates prefill, decode, and storage resources while providing high-performance transfer, distributed KV-cache storage, and fault-tolerant expert-parallel communication.

★ 6.1K⑂ 1KC++
Apache-2.0Q98
Inference, Deployment & Runtime

GPUStack

gpustack/gpustack

GPUStack is an open-source GPU cluster manager for serving AI models and provisioning on-demand, SSH-accessible GPU instances. It orchestrates inference engines including vLLM, SGLang, and TensorRT-LLM across on-premises, Kubernetes, and cloud environments.

★ 5.4K⑂ 604Python
Apache-2.0Q98
Inference, Deployment & Runtime

llm-d: Distributed LLM Inference on Kubernetes

llm-d/llm-d

llm-d is an open-source serving stack that adds distributed orchestration, routing, cache management, autoscaling, and batch processing around model servers such as vLLM and SGLang. It targets high-scale production inference on Kubernetes and modern hardware accelerators.

★ 4K⑂ 649Shell
Apache-2.0Q98
Agents & Multi-Agent

LangWatch

langwatch/langwatch

An open-core platform for evaluating, testing, tracing, and monitoring LLM applications and AI agents before release and in production.

★ 3.5K⑂ 345TypeScript
Apache-2.0Q98
Inference, Deployment & Runtime

vLLM Ascend

vllm-project/vllm-ascend

A community-maintained hardware plugin that enables vLLM to run model inference and serving workloads on supported Ascend NPUs.

★ 2.7K⑂ 2.1KC++
Apache-2.0Q98

Page 6 / 11 · 130 projects

Recently updated

qwen-code — Command-line coding agentQwenLM/qwen-code★ 27.1K SBproxysoapbucket/sbproxy★ 49 XERJxerj-org/xerj★ 1.4K PwrAgentpwrdrvr/pwragent★ 29 OpenGenicloudgeni-ai/opengeni★ 56 NemoClaw — Agent runtime and deployment toolingNVIDIA/NemoClaw★ 22.2K

Most starred

Hermes Agentnousresearch/hermes-agent★ 227.1K markitdown — Document conversion and extractionmicrosoft/markitdown★ 174.3K skills — Reusable agent skills and workflowsanthropics/skills★ 170.1K Hugging Face Transformershuggingface/transformers★ 163.3K Firecrawlfirecrawl/firecrawl★ 161.1K LangChainlangchain-ai/langchain★ 143.6K