ToolAI.io · GitHub Channel

Trending AI Projects on GitHub

Track active open-source AI repositories across LLM, agent, MCP, RAG and coding topics, with stars and verified project facts.

Public repository data

Project index

130 projects
Screenshot of memra
Inference, Deployment & Runtime

memra

avifenesh/memra

A from-scratch LLM inference engine built in Rust and CUDA, specifically optimized for RTX 5090 (sm_120a) and H100 (sm_90a) architectures, delivering exactness-gated performance without relying on frameworks like ggml.

★ 312⑂ 35Rust
MITQ92
Screenshot of 1flowbase
Agents & Multi-Agent

1flowbase

taichuy/1flowbase

An open-source AI gateway that allows local agent clients to publish fusion-style multi-model workflows as OpenAI and Claude-compatible virtual models, providing full observability into traces, tokens, latency, and costs.

★ 257⑂ 19Rust
Apache-2.0Q89
Screenshot of SBproxy
MCP & Tool Calling

SBproxy

soapbucket/sbproxy

An open-source, self-hosted Enterprise AI Gateway and LLM proxy written in Rust, providing an OpenAI-compatible API for over 60 providers and supporting local model hosting, MCP tool federation, and robust traffic governance.

★ 49⑂ 1Rust
Apache-2.0Q86
Screenshot of TensorRT-LLM — Optimized LLM inference
Inference, Deployment & Runtime

TensorRT-LLM — Optimized LLM inference

NVIDIA/TensorRT-LLM

TensorRT LLM provides users with an easy-to-use Python API to define Large Language Models (LLMs) and supports state-of-the-art optimizations to perform inference efficiently on NVIDIA GPUs. TensorRT LLM also contains components to create Python and C++ runtimes that orchestrate the inference execution in a performant way

★ 14.4K⑂ 2.7KC++
NOASSERTIONQ82
Agents & Multi-Agent

dstack

dstackai/dstack

A vendor-agnostic control plane for provisioning and orchestrating training, inference, development, and agentic workloads across GPU clouds, Kubernetes, and on-premises infrastructure.

★ 2.2K⑂ 248Python
MPL-2.0Q98
LLM & Foundation Models

Bionic GPT

bionic-gpt/bionic-gpt

Bionic is a Rust-based, on-premise alternative to ChatGPT designed for private generative AI deployments. It provides a familiar chat interface, local or remote model access, team controls, retrieval-augmented assistants, data integrations, observability, and Kubernetes-oriented scaling.

★ 2.4K⑂ 239Rust
NOASSERTIONQ98
Screenshot of DotCraft
Agents & Multi-Agent

DotCraft

dotharness/dotcraft

An open-source, self-hosted, project-scoped AI agent runtime that provides persistent sessions, memory, background work, and automations shared across Desktop, CLI, bots, and applications.

★ 385⑂ 16C#
Apache-2.0Q88
Inference, Deployment & Runtime

vLLM Ascend

vllm-project/vllm-ascend

A community-maintained hardware plugin that enables vLLM to run model inference and serving workloads on supported Ascend NPUs.

★ 2.7K⑂ 2.1KC++
Apache-2.0Q98
Screenshot of comfyui-mcp
MCP & Tool Calling

comfyui-mcp

artokun/comfyui-mcp

A local-first, agent-native control plane for ComfyUI that provides an MCP server and sidebar agent to generate images, video, and audio, author and run workflows, and edit live graphs using natural language across any LLM.

★ 597⑂ 93TypeScript
MITQ92
Screenshot of Infino
RAG & Knowledge Systems

Infino

infino-ai/infino

Infino is a fast retrieval engine that executes SQL, full-text search, and vector search over a single copy of data stored natively as Parquet on object storage.

★ 67⑂ 18Rust
Apache-2.0Q89

Page 2 / 11 · 130 projects

Recently updated

codex — Coding agent and developer workflowsopenai/codex★ 106.5K NemoClaw — Agent runtime and deployment toolingNVIDIA/NemoClaw★ 22.2K XERJxerj-org/xerj★ 1.4K LangWatchlangwatch/langwatch★ 3.5K OrchestKityonatangross/orchestkit★ 222 hal0hal0ai/hal0★ 67

Most starred

Hermes Agentnousresearch/hermes-agent★ 227.1K markitdown — Document conversion and extractionmicrosoft/markitdown★ 174.2K skills — Reusable agent skills and workflowsanthropics/skills★ 169.9K Hugging Face Transformershuggingface/transformers★ 163.3K Firecrawlfirecrawl/firecrawl★ 161.1K LangChainlangchain-ai/langchain★ 143.6K