ToolAI.io · GitHub Channel

AI Project Directory on GitHub

Browse the complete ToolAI directory of open-source AI projects on GitHub, organized by topic, language and license.

Public repository data

Project index

130 projects
Screenshot of DeepChat - Open-Source Local-First AI Agent Desktop Client
MCP & Tool Calling

DeepChat - Open-Source Local-First AI Agent Desktop Client

thinkinaixyz/deepchat

DeepChat is an open-source, local-first AI agent desktop client built in TypeScript and Electron. It integrates cloud LLMs, local models, MCP services, installable Skills, ACP agents, and remote control for messaging apps, following the Tape.systems philosophy to keep agent sessions recoverable and inspectable.

★ 6.2K⑂ 712TypeScript
Apache-2.0Q98
Inference, Deployment & Runtime

Mooncake: KVCache-Centric Infrastructure for Distributed LLM Serving

kvcache-ai/mooncake

Mooncake is a C++ infrastructure project for large-scale LLM inference and training. It separates prefill, decode, and storage resources while providing high-performance transfer, distributed KV-cache storage, and fault-tolerant expert-parallel communication.

★ 6.1K⑂ 1KC++
Apache-2.0Q98
Screenshot of Potpie
Agents & Multi-Agent

Potpie

potpie-ai/potpie

Potpie transforms a codebase and software development lifecycle into a living context graph for AI agents, indexing code, structure, decisions, source history, and team knowledge to enable context-aware coding, planning, and debugging.

★ 5.5K⑂ 642Python
Apache-2.0Q98
Inference, Deployment & Runtime

GPUStack

gpustack/gpustack

GPUStack is an open-source GPU cluster manager for serving AI models and provisioning on-demand, SSH-accessible GPU instances. It orchestrates inference engines including vLLM, SGLang, and TensorRT-LLM across on-premises, Kubernetes, and cloud environments.

★ 5.4K⑂ 604Python
Apache-2.0Q98
Screenshot of Lemonade: Local AI Server for GPU and NPU Inference
MCP & Tool Calling

Lemonade: Local AI Server for GPU and NPU Inference

lemonade-sdk/lemonade

Lemonade is an open-source local AI server that enables users to run optimized Large Language Models (LLMs), speech, and image generation models directly on their own GPUs and NPUs, providing a free and private alternative to cloud APIs.

★ 5.2K⑂ 436C++
Apache-2.0Q98
Agents & Multi-Agent

Embabel Agent Framework

embabel/embabel-agent

Embabel is an open-source JVM framework for building strongly typed agentic applications that combine LLM interactions, regular code, tools, and domain models. It supports dynamic planning and can be used from Kotlin or Java.

★ 4.3K⑂ 417Kotlin
Apache-2.0Q98
Screenshot of ha-mcp: The Unofficial Home Assistant MCP Server
MCP & Tool Calling

ha-mcp: The Unofficial Home Assistant MCP Server

homeassistant-ai/ha-mcp

A comprehensive Model Context Protocol (MCP) server enabling AI assistants to interact with, configure, build, and debug Home Assistant smart home setups using natural language.

★ 4.3K⑂ 177Python
MITQ98
Screenshot of CSGHub: Open-Source LLM Asset Management Platform
Agents & Multi-Agent

CSGHub: Open-Source LLM Asset Management Platform

opencsgs/csghub

CSGHub is an open-source, on-premise platform for managing the full lifecycle of Large Language Model assets, including models, datasets, spaces, and code, offering functionality comparable to a private Hugging Face.

★ 4.2K⑂ 525Vue
Apache-2.0Q98
Screenshot of OpenConnector
MCP & Tool Calling

OpenConnector

oomol-lab/open-connector

An open-source authentication gateway that connects over 1,000 SaaS providers to AI agents through SDK, CLI, MCP, HTTP, and OpenAPI interfaces.

★ 4.1K⑂ 311TypeScript
Apache-2.0Q98
Screenshot of OpenAgents
Agents & Multi-Agent

OpenAgents

openagents-org/openagents

An open-source platform providing a collaborative operating system for AI agents, enabling unified workspace management, multi-agent coordination, and network integration without vendor lock-in.

★ 4K⑂ 403TypeScript
Apache-2.0Q98
Inference, Deployment & Runtime

llm-d: Distributed LLM Inference on Kubernetes

llm-d/llm-d

llm-d is an open-source serving stack that adds distributed orchestration, routing, cache management, autoscaling, and batch processing around model servers such as vLLM and SGLang. It targets high-scale production inference on Kubernetes and modern hardware accelerators.

★ 4K⑂ 649Shell
Apache-2.0Q98

Page 6 / 11 · 130 projects

Recently updated

qwen-code — Command-line coding agentQwenLM/qwen-code★ 27.1K SBproxysoapbucket/sbproxy★ 49 XERJxerj-org/xerj★ 1.4K PwrAgentpwrdrvr/pwragent★ 29 OpenGenicloudgeni-ai/opengeni★ 56 NemoClaw — Agent runtime and deployment toolingNVIDIA/NemoClaw★ 22.2K

Most starred

Hermes Agentnousresearch/hermes-agent★ 227.1K markitdown — Document conversion and extractionmicrosoft/markitdown★ 174.3K skills — Reusable agent skills and workflowsanthropics/skills★ 170.1K Hugging Face Transformershuggingface/transformers★ 163.3K Firecrawlfirecrawl/firecrawl★ 161.1K LangChainlangchain-ai/langchain★ 143.6K