ToolAI.io · Kênh GitHub

Các dự án AI mã nguồn mở trên GitHub

Chỉ mục dựa trên dữ kiện về các dự án phát triển LLM, tác nhân, MCP, RAG và AI, kèm giấy phép, thiết lập, tải xuống, ảnh chụp màn hình và tài nguyên liên quan cho từng mục.

Dữ liệu kho lưu trữ công khai

Mục lục dự án

24 Dự án
Tác nhân & Đa tác nhân

dstack

dstackai/dstack

A vendor-agnostic control plane for provisioning and orchestrating training, inference, development, and agentic workloads across GPU clouds, Kubernetes, and on-premises infrastructure.

★ 2,2K⑂ 248Python
MPL-2.0Q98
Suy luận, Triển khai & Thời gian chạy

any-llm: A Unified Python Interface for LLM Providers

mozilla-ai/any-llm

any-llm is a Python SDK for communicating with multiple LLM providers through one interface. It supports switching providers and models with minimal code changes while using official provider SDKs when available.

★ 2,2K⑂ 211Python
Apache-2.0Q98
Suy luận, Triển khai & Thời gian chạy

xLLM: High-Performance Inference Engine for Diverse AI Accelerators

xllm-ai/xllm

xLLM is a C++ inference framework for LLM, VLM, DiT and REC models. It targets high-throughput, low-latency deployment on several AI accelerator families and separates service-layer scheduling and availability from engine-layer computation.

★ 1,5K⑂ 279C++
Apache-2.0Q98
Ảnh chụp màn hình của memra
Suy luận, Triển khai & Thời gian chạy

memra

avifenesh/memra

A from-scratch LLM inference engine built in Rust and CUDA, specifically optimized for RTX 5090 (sm_120a) and H100 (sm_90a) architectures, delivering exactness-gated performance without relying on frameworks like ggml.

★ 312⑂ 35Rust
MITQ92
Ảnh chụp màn hình của hal0
Suy luận, Triển khai & Thời gian chạy

hal0

hal0ai/hal0

An open-source, self-hosted home AI inference platform designed to turn a Linux box into an OpenAI-compatible inference appliance, with native optimization for AMD Strix Halo hardware.

★ 67⑂ 7Python
Apache-2.0Q89
Ảnh chụp màn hình của NemoClaw — Agent runtime and deployment tooling
Tác nhân & Đa tác nhân

NemoClaw — Agent runtime and deployment tooling

NVIDIA/NemoClaw

Run agents like Hermes, LangChain Deep Agents, and OpenClaw more securely inside NVIDIA OpenShell with managed inference

★ 22,2K⑂ 3KPython
Apache-2.0Q86
Ảnh chụp màn hình của tiktoken — Fast tokenization
Suy luận, Triển khai & Thời gian chạy

tiktoken — Fast tokenization

openai/tiktoken

tiktoken is a fast BPE tokeniser for use with OpenAI's models

★ 19K⑂ 1,6KRust
MITQ86
Ảnh chụp màn hình của Qwen3-Coder — Code-focused language models
LLM & Mô hình nền tảng

Qwen3-Coder — Code-focused language models

QwenLM/Qwen3-Coder

Qwen3-Coder is the code version of Qwen3, the large language model series developed by Qwen team

★ 16,8K⑂ 1,2KPython
Không phát hiện giấy phépQ82
Ảnh chụp màn hình của TensorRT-LLM — Optimized LLM inference
Suy luận, Triển khai & Thời gian chạy

TensorRT-LLM — Optimized LLM inference

NVIDIA/TensorRT-LLM

TensorRT LLM provides users with an easy-to-use Python API to define Large Language Models (LLMs) and supports state-of-the-art optimizations to perform inference efficiently on NVIDIA GPUs. TensorRT LLM also contains components to create Python and C++ runtimes that orchestrate the inference execution in a performant way

★ 14,4K⑂ 2,7KC++
NOASSERTIONQ82

Trang 2 / 2 · 24 dự án

Đã cập nhật gần đây

codex — Coding agent and developer workflowsopenai/codex★ 106,5K NemoClaw — Agent runtime and deployment toolingNVIDIA/NemoClaw★ 22,2K XERJxerj-org/xerj★ 1,4K LangWatchlangwatch/langwatch★ 3,5K OrchestKityonatangross/orchestkit★ 222 hal0hal0ai/hal0★ 67

Được gắn sao nhiều nhất

Hermes Agentnousresearch/hermes-agent★ 227,1K markitdown — Document conversion and extractionmicrosoft/markitdown★ 174,2K skills — Reusable agent skills and workflowsanthropics/skills★ 169,9K Hugging Face Transformershuggingface/transformers★ 163,3K Firecrawlfirecrawl/firecrawl★ 161,1K LangChainlangchain-ai/langchain★ 143,6K