ToolAI.io · GitHub Channel

Open-source AI Projects on GitHub

A fact-based index of LLM, agent, MCP, RAG and AI development projects, with license, setup, download, screenshots and related resources for each entry.

Public repository data

Project index

75 projects
Inference, Deployment & Runtime

vLLM Ascend

vllm-project/vllm-ascend

A community-maintained hardware plugin that enables vLLM to run model inference and serving workloads on supported Ascend NPUs.

★ 2.7K⑂ 2.1KC++
Apache-2.0Q98
LLM & Foundation Models

Bionic GPT

bionic-gpt/bionic-gpt

Bionic is a Rust-based, on-premise alternative to ChatGPT designed for private generative AI deployments. It provides a familiar chat interface, local or remote model access, team controls, retrieval-augmented assistants, data integrations, observability, and Kubernetes-oriented scaling.

★ 2.4K⑂ 239Rust
NOASSERTIONQ98
Inference, Deployment & Runtime

any-llm: A Unified Python Interface for LLM Providers

mozilla-ai/any-llm

any-llm is a Python SDK for communicating with multiple LLM providers through one interface. It supports switching providers and models with minimal code changes while using official provider SDKs when available.

★ 2.2K⑂ 211Python
Apache-2.0Q98
Agents & Multi-Agent

Future AGI

future-agi/future-agi

An open-source, self-hostable platform for evaluating, tracing, simulating, protecting, routing, and optimizing LLM and AI-agent applications.

★ 1.7K⑂ 497Python
Apache-2.0Q98
Inference, Deployment & Runtime

xLLM: High-Performance Inference Engine for Diverse AI Accelerators

xllm-ai/xllm

xLLM is a C++ inference framework for LLM, VLM, DiT and REC models. It targets high-throughput, low-latency deployment on several AI accelerator families and separates service-layer scheduling and availability from engine-layer computation.

★ 1.5K⑂ 279C++
Apache-2.0Q98
Screenshot of Askimo
MCP & Tool Calling

Askimo

askimo-ai/askimo

A native desktop AI client for chat, local RAG, multi-step AI workflows (Plans), and agent skills, supporting multiple cloud and local LLM providers while keeping user files strictly on the machine.

★ 334⑂ 71Kotlin
AGPL-3.0Q94
Screenshot of NeuroLink
MCP & Tool Calling

NeuroLink

juspay/neurolink

A TypeScript integration platform providing a unified API for 30+ AI providers and 100+ models, enabling provider swapping, multi-modal voice processing, RAG, memory, and MCP-native tool integration.

★ 121⑂ 124TypeScript
MITQ93
Screenshot of comfyui-mcp
MCP & Tool Calling

comfyui-mcp

artokun/comfyui-mcp

A local-first, agent-native control plane for ComfyUI that provides an MCP server and sidebar agent to generate images, video, and audio, author and run workflows, and edit live graphs using natural language across any LLM.

★ 597⑂ 93TypeScript
MITQ92
Screenshot of models
Agents & Multi-Agent

models

reyamira/models

A TUI and CLI tool for browsing AI models, benchmarks, coding agents, and provider statuses.

★ 492⑂ 19Rust
MITQ92
Screenshot of memra
Inference, Deployment & Runtime

memra

avifenesh/memra

A from-scratch LLM inference engine built in Rust and CUDA, specifically optimized for RTX 5090 (sm_120a) and H100 (sm_90a) architectures, delivering exactness-gated performance without relying on frameworks like ggml.

★ 312⑂ 35Rust
MITQ92
Screenshot of LLMix
Agents & Multi-Agent

LLMix

sno-ai/llmix

A production LLM call layer for AI agents and tools that wraps existing provider SDKs with config-driven model presets, caching, resilience patterns, and key rotation across Python, TypeScript, and Rust.

★ 131⑂ 28Python
Apache-2.0Q91
Screenshot of 1flowbase
Agents & Multi-Agent

1flowbase

taichuy/1flowbase

An open-source AI gateway that allows local agent clients to publish fusion-style multi-model workflows as OpenAI and Claude-compatible virtual models, providing full observability into traces, tokens, latency, and costs.

★ 257⑂ 19Rust
Apache-2.0Q89

Page 5 / 7 · 75 projects

Recently updated

codex — Coding agent and developer workflowsopenai/codex★ 106.5K NemoClaw — Agent runtime and deployment toolingNVIDIA/NemoClaw★ 22.2K XERJxerj-org/xerj★ 1.4K LangWatchlangwatch/langwatch★ 3.5K OrchestKityonatangross/orchestkit★ 222 hal0hal0ai/hal0★ 67

Most starred

Hermes Agentnousresearch/hermes-agent★ 227.1K markitdown — Document conversion and extractionmicrosoft/markitdown★ 174.2K skills — Reusable agent skills and workflowsanthropics/skills★ 169.9K Hugging Face Transformershuggingface/transformers★ 163.3K Firecrawlfirecrawl/firecrawl★ 161.1K LangChainlangchain-ai/langchain★ 143.6K