ToolAI.io · GitHub 채널

GitHub에서 인기 있는 AI 프로젝트

LLM, 에이전트, MCP, RAG 및 코딩 분야의 활발한 오픈 소스 AI 리포지토리를 추적하고, 스타 수와 검증된 프로젝트 정보를 제공합니다.

공개 리포지토리 데이터

프로젝트 색인

130 프로젝트
memra 스크린샷
추론, 배포 및 런타임

memra

avifenesh/memra

A from-scratch LLM inference engine built in Rust and CUDA, specifically optimized for RTX 5090 (sm_120a) and H100 (sm_90a) architectures, delivering exactness-gated performance without relying on frameworks like ggml.

★ 312⑂ 35Rust
MITQ92
1flowbase 스크린샷
에이전트 및 멀티 에이전트

1flowbase

taichuy/1flowbase

An open-source AI gateway that allows local agent clients to publish fusion-style multi-model workflows as OpenAI and Claude-compatible virtual models, providing full observability into traces, tokens, latency, and costs.

★ 257⑂ 19Rust
Apache-2.0Q89
SBproxy 스크린샷
MCP 및 도구 호출

SBproxy

soapbucket/sbproxy

An open-source, self-hosted Enterprise AI Gateway and LLM proxy written in Rust, providing an OpenAI-compatible API for over 60 providers and supporting local model hosting, MCP tool federation, and robust traffic governance.

★ 49⑂ 1Rust
Apache-2.0Q86
TensorRT-LLM — Optimized LLM inference 스크린샷
추론, 배포 및 런타임

TensorRT-LLM — Optimized LLM inference

NVIDIA/TensorRT-LLM

TensorRT LLM provides users with an easy-to-use Python API to define Large Language Models (LLMs) and supports state-of-the-art optimizations to perform inference efficiently on NVIDIA GPUs. TensorRT LLM also contains components to create Python and C++ runtimes that orchestrate the inference execution in a performant way

★ 14.4K⑂ 2.7KC++
NOASSERTIONQ82
에이전트 및 멀티 에이전트

dstack

dstackai/dstack

A vendor-agnostic control plane for provisioning and orchestrating training, inference, development, and agentic workloads across GPU clouds, Kubernetes, and on-premises infrastructure.

★ 2.2K⑂ 248Python
MPL-2.0Q98
LLM 및 파운데이션 모델

Bionic GPT

bionic-gpt/bionic-gpt

Bionic is a Rust-based, on-premise alternative to ChatGPT designed for private generative AI deployments. It provides a familiar chat interface, local or remote model access, team controls, retrieval-augmented assistants, data integrations, observability, and Kubernetes-oriented scaling.

★ 2.4K⑂ 239Rust
NOASSERTIONQ98
DotCraft 스크린샷
에이전트 및 멀티 에이전트

DotCraft

dotharness/dotcraft

An open-source, self-hosted, project-scoped AI agent runtime that provides persistent sessions, memory, background work, and automations shared across Desktop, CLI, bots, and applications.

★ 385⑂ 16C#
Apache-2.0Q88
추론, 배포 및 런타임

vLLM Ascend

vllm-project/vllm-ascend

A community-maintained hardware plugin that enables vLLM to run model inference and serving workloads on supported Ascend NPUs.

★ 2.7K⑂ 2.1KC++
Apache-2.0Q98
comfyui-mcp 스크린샷
MCP 및 도구 호출

comfyui-mcp

artokun/comfyui-mcp

A local-first, agent-native control plane for ComfyUI that provides an MCP server and sidebar agent to generate images, video, and audio, author and run workflows, and edit live graphs using natural language across any LLM.

★ 597⑂ 93TypeScript
MITQ92
Infino 스크린샷
RAG 및 지식 시스템

Infino

infino-ai/infino

Infino is a fast retrieval engine that executes SQL, full-text search, and vector search over a single copy of data stored natively as Parquet on object storage.

★ 67⑂ 18Rust
Apache-2.0Q89

2 / 11페이지 · 프로젝트 130개

최근 업데이트

codex — Coding agent and developer workflowsopenai/codex★ 106.5K NemoClaw — Agent runtime and deployment toolingNVIDIA/NemoClaw★ 22.2K XERJxerj-org/xerj★ 1.4K LangWatchlangwatch/langwatch★ 3.5K OrchestKityonatangross/orchestkit★ 222 hal0hal0ai/hal0★ 67

별이 가장 많은

Hermes Agentnousresearch/hermes-agent★ 227.1K markitdown — Document conversion and extractionmicrosoft/markitdown★ 174.2K skills — Reusable agent skills and workflowsanthropics/skills★ 169.9K Hugging Face Transformershuggingface/transformers★ 163.3K Firecrawlfirecrawl/firecrawl★ 161.1K LangChainlangchain-ai/langchain★ 143.6K