Agen & Multi-Agen
opencsgs/csghub
CSGHub is an open-source, on-premise platform for managing the full lifecycle of Large Language Model assets, including models, datasets, spaces, and code, offering functionality comparable to a private Hugging Face.
★ 4,2K⑂ 525Vue
Apache-2.0Q98
MCP & Pemanggilan Tool
lemonade-sdk/lemonade
Lemonade is an open-source local AI server that enables users to run optimized Large Language Models (LLMs), speech, and image generation models directly on their own GPUs and NPUs, providing a free and private alternative to cloud APIs.
★ 5,2K⑂ 436C++
Apache-2.0Q98
Inferensi, Deployment & Runtime
openvinotoolkit/openvino
Open-source toolkit for optimizing and deploying AI inference across edge-to-cloud environments.
★ 10,6K⑂ 3,3KC++
Apache-2.0Q98
Inferensi, Deployment & Runtime
kvcache-ai/mooncake
Mooncake is a C++ infrastructure project for large-scale LLM inference and training. It separates prefill, decode, and storage resources while providing high-performance transfer, distributed KV-cache storage, and fault-tolerant expert-parallel communication.
★ 6,1K⑂ 1KC++
Apache-2.0Q98
Inferensi, Deployment & Runtime
llm-d/llm-d
llm-d is an open-source serving stack that adds distributed orchestration, routing, cache management, autoscaling, and batch processing around model servers such as vLLM and SGLang. It targets high-scale production inference on Kubernetes and modern hardware accelerators.
★ 4K⑂ 649Shell
Apache-2.0Q98
Inferensi, Deployment & Runtime
gpustack/gpustack
GPUStack is an open-source GPU cluster manager for serving AI models and provisioning on-demand, SSH-accessible GPU instances. It orchestrates inference engines including vLLM, SGLang, and TensorRT-LLM across on-premises, Kubernetes, and cloud environments.
★ 5,4K⑂ 604Python
Apache-2.0Q98
LLM & Model Fondasi
openai/gpt-oss
gpt-oss-120b and gpt-oss-20b are two open-weight language models by OpenAI
★ 20,3K⑂ 2,1KPython
Apache-2.0Q86
LLM & Model Fondasi
QwenLM/Qwen3-Coder
Qwen3-Coder is the code version of Qwen3, the large language model series developed by Qwen team
★ 16,8K⑂ 1,2KPython
Lisensi tidak terdeteksiQ82
LLM & Model Fondasi
QwenLM/Qwen3
Qwen3 is the large language model series developed by Qwen team, Alibaba Cloud
★ 27,5K⑂ 2KPython
Lisensi tidak terdeteksiQ82
LLM & Model Fondasi
deepseek-ai/DeepSeek-Coder
DeepSeek Coder: Let the Code Write Itself
★ 24,2K⑂ 2,9KPython
MITQ86
LLM & Model Fondasi
Unggulan
deepseek-ai/DeepSeek-V3
An open-source project focused on large language model research and inference.
★ 104,3K⑂ 16,7KPython
MITQ86
LLM & Model Fondasi
Unggulan
deepseek-ai/DeepSeek-R1
An open-source project focused on reasoning model research and evaluation.
★ 92K⑂ 11,7KPython
MITQ86