ToolAI.io · Канал GitHub

Популярные AI-проекты на GitHub

Отслеживайте активные AI-репозитории с открытым исходным кодом по темам LLM, агентов, MCP, RAG и программирования, с количеством звёзд и проверенными сведениями о проектах.

Данные публичного репозитория

Индекс проекта

253 Проекты
MCP и вызов инструментов

Rapid-MLX

raullenchai/rapid-mlx

The fastest local AI engine for Apple Silicon. 4.2x faster than Ollama, 0.08s cached TTFT, 100% tool calling. 17 tool parsers, prompt cache, reasoning separation, cloud routing. Drop-in OpenAI replacement. Works with Claude Code, Cursor, Aider.

★ 3,5K⑂ 401Python
NOASSERTIONQ96
Инференс, развёртывание и среда выполнения

llmgateway

theopenco/llmgateway

Route, manage, and analyze your LLM requests across multiple providers with a unified API interface.

★ 1,6K⑂ 172TypeScript
NOASSERTIONQ98
Инференс, развёртывание и среда выполнения

sglang

sgl-project/sglang

SGLang is a high-performance serving framework for large language models and multimodal models.

★ 32,2K⑂ 8,1KPython
Apache-2.0Q98
Дообучение, обучение и данные

DeepSpeed

deepspeedai/deepspeed

DeepSpeed is a deep learning optimization library that makes distributed training and inference easy, efficient, and effective.

★ 43K⑂ 4,9KPython
Apache-2.0Q94
Агенты и мультиагенты

ctx

ctxrs/ctx

Search the coding agent history already on your machine

★ 1K⑂ 61Rust
Apache-2.0Q90
Инференс, развёртывание и среда выполнения

vllm

vllm-project/vllm

A high-throughput and memory-efficient inference and serving engine for LLMs

★ 89,5K⑂ 21KPython
Apache-2.0Q98
Агенты и мультиагенты

CodeWhale

hmbown/codewhale

Open-source, community-driven agent harness

★ 40,8K⑂ 3,5KRust
MITQ92
Агенты и мультиагенты

matrixone

matrixorigin/matrixone

AI-native HTAP database with Git-for-Data and built-in vector search, serving as the data and memory backbone for intelligent agents and applications.

★ 1,9K⑂ 307Go
Apache-2.0Q94
Скриншот 1flowbase
Агенты и мультиагенты

1flowbase

taichuy/1flowbase

An open-source AI gateway that allows local agent clients to publish fusion-style multi-model workflows as OpenAI and Claude-compatible virtual models, providing full observability into traces, tokens, latency, and costs.

★ 259⑂ 19Rust
Apache-2.0Q89
Скриншот TensorRT-LLM — Optimized LLM inference
Инференс, развёртывание и среда выполнения

TensorRT-LLM — Optimized LLM inference

NVIDIA/TensorRT-LLM

TensorRT LLM provides users with an easy-to-use Python API to define Large Language Models (LLMs) and supports state-of-the-art optimizations to perform inference efficiently on NVIDIA GPUs. TensorRT LLM also contains components to create Python and C++ runtimes that orchestrate the inference execution in a performant way

★ 14,4K⑂ 2,7KC++
NOASSERTIONQ82
Скриншот hal0
Инференс, развёртывание и среда выполнения

hal0

hal0ai/hal0

An open-source, self-hosted home AI inference platform designed to turn a Linux box into an OpenAI-compatible inference appliance, with native optimization for AMD Strix Halo hardware.

★ 68⑂ 7Python
Apache-2.0Q89

Страница 5 / 22 · 253 проектов

Недавно обновлено

cherry-studioCherryHQ/cherry-studio★ 51,5K siyuansiyuan-note/siyuan★ 46,2K career-opscareer-ops-hq/career-ops★ 70,2K tensorflowtensorflow/tensorflow★ 198,8K streamlitstreamlit/streamlit★ 45,7K pytorchpytorch/pytorch★ 102,8K

Наибольшее число звёзд

ECCaffaan-m/ECC★ 248,8K Hermes Agentnousresearch/hermes-agent★ 227,1K tensorflowtensorflow/tensorflow★ 198,8K AutoGPTSignificant-Gravitas/AutoGPT★ 187,1K ollamaollama/ollama★ 180,2K markitdown — Document conversion and extractionmicrosoft/markitdown★ 174,9K