ToolAI.io · Kênh GitHub

Các dự án AI nổi bật trên GitHub

Theo dõi các kho lưu trữ AI mã nguồn mở đang hoạt động trong các chủ đề LLM, tác nhân, MCP, RAG và lập trình, kèm số sao và thông tin dự án đã xác minh.

Dữ liệu kho lưu trữ công khai

Mục lục dự án

41 Dự án
Suy luận, Triển khai & Thời gian chạy

OpenVINO

openvinotoolkit/openvino

Open-source toolkit for optimizing and deploying AI inference across edge-to-cloud environments.

★ 10,6K⑂ 3,3KC++
Apache-2.0Q98
Suy luận, Triển khai & Thời gian chạy

Mooncake: KVCache-Centric Infrastructure for Distributed LLM Serving

kvcache-ai/mooncake

Mooncake is a C++ infrastructure project for large-scale LLM inference and training. It separates prefill, decode, and storage resources while providing high-performance transfer, distributed KV-cache storage, and fault-tolerant expert-parallel communication.

★ 6,1K⑂ 1KC++
Apache-2.0Q98
Suy luận, Triển khai & Thời gian chạy

llm-d: Distributed LLM Inference on Kubernetes

llm-d/llm-d

llm-d is an open-source serving stack that adds distributed orchestration, routing, cache management, autoscaling, and batch processing around model servers such as vLLM and SGLang. It targets high-scale production inference on Kubernetes and modern hardware accelerators.

★ 4K⑂ 649Shell
Apache-2.0Q98
Suy luận, Triển khai & Thời gian chạy

GPUStack

gpustack/gpustack

GPUStack is an open-source GPU cluster manager for serving AI models and provisioning on-demand, SSH-accessible GPU instances. It orchestrates inference engines including vLLM, SGLang, and TensorRT-LLM across on-premises, Kubernetes, and cloud environments.

★ 5,4K⑂ 604Python
Apache-2.0Q98
Suy luận, Triển khai & Thời gian chạy

Model-Optimizer

NVIDIA/Model-Optimizer

A unified library of SOTA model optimization techniques like quantization, distillation, pruning, neural architecture search, speculative decoding, etc. It compresses deep learning models for downstream deployment frameworks like TensorRT-LLM, TensorRT, vLLM, etc. to optimize inference speed.

★ 3,3K⑂ 513Python
Apache-2.0Q86
Suy luận, Triển khai & Thời gian chạy

FastDeploy

PaddlePaddle/FastDeploy

High-performance Inference and Deployment Toolkit for LLMs and VLMs based on PaddlePaddle

★ 3,7K⑂ 755Python
Apache-2.0Q98
Suy luận, Triển khai & Thời gian chạy

TensorRT

NVIDIA/TensorRT

NVIDIA® TensorRT™ is an SDK for high-performance deep learning inference on NVIDIA GPUs. This repository contains the open source components of TensorRT.

★ 13,2K⑂ 2,4KC++
Apache-2.0Q94
Suy luận, Triển khai & Thời gian chạy

checkpoint-engine

MoonshotAI/checkpoint-engine

Checkpoint-engine is a simple middleware to update model weights in LLM inference engines

★ 982⑂ 100Python
MITQ83
Thị giác máy tính

sam2

facebookresearch/sam2

The repository provides code for running inference with the Meta Segment Anything Model 2 (SAM 2), links for downloading the trained model checkpoints, and example notebooks that show how to use the model.

★ 19,6K⑂ 2,5KJupyter Notebook
Apache-2.0Q79
Suy luận, Triển khai & Thời gian chạy

Paddle-Lite

PaddlePaddle/Paddle-Lite

PaddlePaddle High Performance Deep Learning Inference Engine for Mobile and Edge (飞桨高性能深度学习端侧推理引擎)

★ 7,3K⑂ 1,6KC++
Apache-2.0Q91
Robot và Edge AI

edgeai-for-beginners

microsoft/edgeai-for-beginners

This course is designed to guide beginners through the exciting world of Edge AI, covering fundamental concepts, popular models, inference techniques, device-specific applications, model optimization, and the development of intelligent Edge AI agents.

★ 1,6K⑂ 362Jupyter Notebook
MITQ83

Trang 3 / 4 · 41 dự án

Đã cập nhật gần đây

cherry-studioCherryHQ/cherry-studio★ 51,5K siyuansiyuan-note/siyuan★ 46,2K career-opscareer-ops-hq/career-ops★ 70,2K tensorflowtensorflow/tensorflow★ 198,8K streamlitstreamlit/streamlit★ 45,7K pytorchpytorch/pytorch★ 102,8K

Được gắn sao nhiều nhất

ECCaffaan-m/ECC★ 248,8K Hermes Agentnousresearch/hermes-agent★ 227,1K tensorflowtensorflow/tensorflow★ 198,8K AutoGPTSignificant-Gravitas/AutoGPT★ 187,1K ollamaollama/ollama★ 180,2K markitdown — Document conversion and extractionmicrosoft/markitdown★ 174,9K