Inference, Deployment & Runtime

vllm

vllm-project/vllm

A high-throughput and memory-efficient inference and serving engine for LLMs

★ 89.5KStars
⑂ 21KForks
6832Open issues
PythonLanguage
Apache-2.0License
Q98Editorial score

Overview

A high-throughput and memory-efficient inference and serving engine for LLMs

Requirements, installation and quick start

Installation requirements are not stated in the repository metadata.

Usage

Usage is not stated in the repository metadata.

Model compatibility and use cases

Model compatibility is not stated in the repository metadata.

License and risk notes

Apache-2.0

ToolAI metadata-only listing: repository metadata is public; editorial content and manual verification are still pending.

sglang

sgl-project/sglang

★ 32.2KPython

OpenVINO

openvinotoolkit/openvino

★ 10.6KC++

GPUStack

gpustack/gpustack

★ 5.4KPython

FastDeploy

PaddlePaddle/FastDeploy

★ 3.7KPython