Inference, Deployment & Runtime

vllm

vllm-project/vllm

A high-throughput and memory-efficient inference and serving engine for LLMs

★ 89.4KStars
⑂ 20.9KForks
6746Open issues
PythonLanguage
Apache-2.0License
Q@project.QualityScoreEditorial score

Overview

A high-throughput and memory-efficient inference and serving engine for LLMs

Requirements, installation and quick start

Installation requirements are not stated in the repository metadata.

Usage

Usage is not stated in the repository metadata.

Model compatibility and use cases

Model compatibility is not stated in the repository metadata.

License and risk notes

Apache-2.0

ToolAI metadata-only listing: repository metadata is public; editorial content and manual verification are still pending.

sglang

sgl-project/sglang

★ 32.1KPython

OpenVINO

openvinotoolkit/openvino

★ 10.6KC++

GPUStack

gpustack/gpustack

★ 5.4KPython