Inferenz, Bereitstellung & Laufzeit

vllm

vllm-project/vllm

A high-throughput and memory-efficient inference and serving engine for LLMs

★ 89,4KSterne
⑂ 20,9KForks
6746Offene Issues
PythonSprache
Apache-2.0Lizenz
Q@project.QualityScoreRedaktionelle Bewertung

Übersicht

A high-throughput and memory-efficient inference and serving engine for LLMs

Voraussetzungen, Installation und Schnellstart

Installationsanforderungen sind in den Repository-Metadaten nicht angegeben.

Nutzung

Die Nutzung ist in den Repository-Metadaten nicht angegeben.

Modellkompatibilität und Anwendungsfälle

Die Modellkompatibilität ist in den Repository-Metadaten nicht angegeben.

Lizenz- und Risikohinweise

Apache-2.0

ToolAI metadata-only listing: repository metadata is public; editorial content and manual verification are still pending.

sglang

sgl-project/sglang

★ 32,1KPython

OpenVINO

openvinotoolkit/openvino

★ 10,6KC++

GPUStack

gpustack/gpustack

★ 5,4KPython