エージェントとマルチエージェント
opencsgs/csghub
CSGHub is an open-source, on-premise platform for managing the full lifecycle of Large Language Model assets, including models, datasets, spaces, and code, offering functionality comparable to a private Hugging Face.
★ 4.2K⑂ 525Vue
Apache-2.0Q98
MCPとツール呼び出し
lemonade-sdk/lemonade
Lemonade is an open-source local AI server that enables users to run optimized Large Language Models (LLMs), speech, and image generation models directly on their own GPUs and NPUs, providing a free and private alternative to cloud APIs.
★ 5.2K⑂ 436C++
Apache-2.0Q98
推論、デプロイ、ランタイム
openvinotoolkit/openvino
Open-source toolkit for optimizing and deploying AI inference across edge-to-cloud environments.
★ 10.6K⑂ 3.3KC++
Apache-2.0Q98
推論、デプロイ、ランタイム
kvcache-ai/mooncake
Mooncake is a C++ infrastructure project for large-scale LLM inference and training. It separates prefill, decode, and storage resources while providing high-performance transfer, distributed KV-cache storage, and fault-tolerant expert-parallel communication.
★ 6.1K⑂ 1KC++
Apache-2.0Q98
推論、デプロイ、ランタイム
llm-d/llm-d
llm-d is an open-source serving stack that adds distributed orchestration, routing, cache management, autoscaling, and batch processing around model servers such as vLLM and SGLang. It targets high-scale production inference on Kubernetes and modern hardware accelerators.
★ 4K⑂ 649Shell
Apache-2.0Q98
推論、デプロイ、ランタイム
gpustack/gpustack
GPUStack is an open-source GPU cluster manager for serving AI models and provisioning on-demand, SSH-accessible GPU instances. It orchestrates inference engines including vLLM, SGLang, and TensorRT-LLM across on-premises, Kubernetes, and cloud environments.
★ 5.4K⑂ 604Python
Apache-2.0Q98
LLMと基盤モデル
openai/gpt-oss
gpt-oss-120b and gpt-oss-20b are two open-weight language models by OpenAI
★ 20.3K⑂ 2.1KPython
Apache-2.0Q86
LLMと基盤モデル
QwenLM/Qwen3-Coder
Qwen3-Coder is the code version of Qwen3, the large language model series developed by Qwen team
★ 16.8K⑂ 1.2KPython
ライセンス未検出Q82
LLMと基盤モデル
QwenLM/Qwen3
Qwen3 is the large language model series developed by Qwen team, Alibaba Cloud
★ 27.5K⑂ 2KPython
ライセンス未検出Q82
LLMと基盤モデル
deepseek-ai/DeepSeek-Coder
DeepSeek Coder: Let the Code Write Itself
★ 24.2K⑂ 2.9KPython
MITQ86
LLMと基盤モデル
注目
deepseek-ai/DeepSeek-V3
An open-source project focused on large language model research and inference.
★ 104.3K⑂ 16.7KPython
MITQ86
LLMと基盤モデル
注目
deepseek-ai/DeepSeek-R1
An open-source project focused on reasoning model research and evaluation.
★ 92K⑂ 11.7KPython
MITQ86