ToolAI.io · GitHub Channel

Open-source AI Projects on GitHub

A fact-based index of LLM, agent, MCP, RAG and AI development projects, with license, setup, download, screenshots and related resources for each entry.

Public repository data

Project index

118 projects
Inference, Deployment & Runtime

ray

ray-project/ray

A distributed computing runtime for machine learning training, tuning and model serving.

★ 43.7K⑂ 8KPython
Apache-2.0Q85
LLM & Foundation Models

MiniMax-M2

MiniMax-AI/MiniMax-M2

MiniMax-M2, a model built for Max coding & agentic workflows.

★ 2.6K⑂ 215
NOASSERTIONQ83
Inference, Deployment & Runtime

checkpoint-engine

MoonshotAI/checkpoint-engine

Checkpoint-engine is a simple middleware to update model weights in LLM inference engines

★ 982⑂ 100Python
MITQ83
Screenshot of evals — Model evaluation framework
Evaluation, Observability & Safety

evals — Model evaluation framework

openai/evals

Evals is a framework for evaluating LLMs and LLM systems, and an open-source registry of benchmarks

★ 19.2K⑂ 3.1KPython
NOASSERTIONQ82
Screenshot of Qwen3-Coder — Code-focused language models
LLM & Foundation Models

Qwen3-Coder — Code-focused language models

QwenLM/Qwen3-Coder

Qwen3-Coder is the code version of Qwen3, the large language model series developed by Qwen team

★ 16.8K⑂ 1.2KPython
License not detectedQ82
Screenshot of TensorRT-LLM — Optimized LLM inference
Inference, Deployment & Runtime

TensorRT-LLM — Optimized LLM inference

NVIDIA/TensorRT-LLM

TensorRT LLM provides users with an easy-to-use Python API to define Large Language Models (LLMs) and supports state-of-the-art optimizations to perform inference efficiently on NVIDIA GPUs. TensorRT LLM also contains components to create Python and C++ runtimes that orchestrate the inference execution in a performant way

★ 14.4K⑂ 2.7KC++
NOASSERTIONQ82
LLM & Foundation Models

PaddleFormers

PaddlePaddle/PaddleFormers

PaddleFormers is an easy-to-use library of pre-trained large language model zoo based on PaddlePaddle.

★ 13K⑂ 2.2KPython
Apache-2.0Q81
Inference, Deployment & Runtime

jan

janhq/jan

A local AI chat application for running models on a personal computer.

★ 44.3K⑂ 3KTypeScript
License not detectedQ80

Page 10 / 10 · 118 projects

Recently updated

cherry-studioCherryHQ/cherry-studio★ 51.5K siyuansiyuan-note/siyuan★ 46.2K career-opscareer-ops-hq/career-ops★ 70.2K tensorflowtensorflow/tensorflow★ 198.8K streamlitstreamlit/streamlit★ 45.7K pytorchpytorch/pytorch★ 102.8K

Most starred

ECCaffaan-m/ECC★ 248.8K Hermes Agentnousresearch/hermes-agent★ 227.1K tensorflowtensorflow/tensorflow★ 198.8K AutoGPTSignificant-Gravitas/AutoGPT★ 187.1K ollamaollama/ollama★ 180.2K markitdown — Document conversion and extractionmicrosoft/markitdown★ 174.9K