ToolAI.io · GitHub चैनल

GitHub पर ट्रेंडिंग AI प्रोजेक्ट्स

LLM, एजेंट, MCP, RAG और कोडिंग विषयों में सक्रिय ओपन-सोर्स AI रिपॉज़िटरीज़ को स्टार्स और सत्यापित प्रोजेक्ट जानकारी के साथ ट्रैक करें।

सार्वजनिक रिपॉज़िटरी का डेटा

प्रोजेक्ट इंडेक्स

253 प्रोजेक्ट्स
MCP और टूल कॉलिंग

Rapid-MLX

raullenchai/rapid-mlx

The fastest local AI engine for Apple Silicon. 4.2x faster than Ollama, 0.08s cached TTFT, 100% tool calling. 17 tool parsers, prompt cache, reasoning separation, cloud routing. Drop-in OpenAI replacement. Works with Claude Code, Cursor, Aider.

★ 3.5K⑂ 401Python
NOASSERTIONQ96
इन्फरेंस, डिप्लॉयमेंट और रनटाइम

llmgateway

theopenco/llmgateway

Route, manage, and analyze your LLM requests across multiple providers with a unified API interface.

★ 1.6K⑂ 172TypeScript
NOASSERTIONQ98
इन्फरेंस, डिप्लॉयमेंट और रनटाइम

sglang

sgl-project/sglang

SGLang is a high-performance serving framework for large language models and multimodal models.

★ 32.2K⑂ 8.1KPython
Apache-2.0Q98
फ़ाइन-ट्यूनिंग, प्रशिक्षण और डेटा

DeepSpeed

deepspeedai/deepspeed

DeepSpeed is a deep learning optimization library that makes distributed training and inference easy, efficient, and effective.

★ 43K⑂ 4.9KPython
Apache-2.0Q94
एजेंट और मल्टी-एजेंट

ctx

ctxrs/ctx

Search the coding agent history already on your machine

★ 1K⑂ 61Rust
Apache-2.0Q90
इन्फरेंस, डिप्लॉयमेंट और रनटाइम

vllm

vllm-project/vllm

A high-throughput and memory-efficient inference and serving engine for LLMs

★ 89.5K⑂ 21KPython
Apache-2.0Q98
एजेंट और मल्टी-एजेंट

CodeWhale

hmbown/codewhale

Open-source, community-driven agent harness

★ 40.8K⑂ 3.5KRust
MITQ92
एजेंट और मल्टी-एजेंट

matrixone

matrixorigin/matrixone

AI-native HTAP database with Git-for-Data and built-in vector search, serving as the data and memory backbone for intelligent agents and applications.

★ 1.9K⑂ 307Go
Apache-2.0Q94
1flowbase का स्क्रीनशॉट
एजेंट और मल्टी-एजेंट

1flowbase

taichuy/1flowbase

An open-source AI gateway that allows local agent clients to publish fusion-style multi-model workflows as OpenAI and Claude-compatible virtual models, providing full observability into traces, tokens, latency, and costs.

★ 259⑂ 19Rust
Apache-2.0Q89
TensorRT-LLM — Optimized LLM inference का स्क्रीनशॉट
इन्फरेंस, डिप्लॉयमेंट और रनटाइम

TensorRT-LLM — Optimized LLM inference

NVIDIA/TensorRT-LLM

TensorRT LLM provides users with an easy-to-use Python API to define Large Language Models (LLMs) and supports state-of-the-art optimizations to perform inference efficiently on NVIDIA GPUs. TensorRT LLM also contains components to create Python and C++ runtimes that orchestrate the inference execution in a performant way

★ 14.4K⑂ 2.7KC++
NOASSERTIONQ82
hal0 का स्क्रीनशॉट
इन्फरेंस, डिप्लॉयमेंट और रनटाइम

hal0

hal0ai/hal0

An open-source, self-hosted home AI inference platform designed to turn a Linux box into an OpenAI-compatible inference appliance, with native optimization for AMD Strix Halo hardware.

★ 68⑂ 7Python
Apache-2.0Q89

पृष्ठ 5 / 22 · 253 प्रोजेक्ट

हाल ही में अपडेट किए गए

cherry-studioCherryHQ/cherry-studio★ 51.5K siyuansiyuan-note/siyuan★ 46.2K career-opscareer-ops-hq/career-ops★ 70.2K tensorflowtensorflow/tensorflow★ 198.8K streamlitstreamlit/streamlit★ 45.7K pytorchpytorch/pytorch★ 102.8K

सर्वाधिक स्टार वाले

ECCaffaan-m/ECC★ 248.8K Hermes Agentnousresearch/hermes-agent★ 227.1K tensorflowtensorflow/tensorflow★ 198.8K AutoGPTSignificant-Gravitas/AutoGPT★ 187.1K ollamaollama/ollama★ 180.2K markitdown — Document conversion and extractionmicrosoft/markitdown★ 174.9K