Inferenz, Bereitstellung & Laufzeit

vllm

vllm-project/vllm

A high-throughput and memory-efficient inference and serving engine for LLMs

★ 89,5KSterne
⑂ 21KForks
6832Offene Issues
PythonSprache
Apache-2.0Lizenz
Q98Redaktionelle Bewertung

Übersicht

A high-throughput and memory-efficient inference and serving engine for LLMs

Voraussetzungen, Installation und Schnellstart

Installationsanforderungen sind in den Repository-Metadaten nicht angegeben.

Nutzung

Die Nutzung ist in den Repository-Metadaten nicht angegeben.

Modellkompatibilität und Anwendungsfälle

Die Modellkompatibilität ist in den Repository-Metadaten nicht angegeben.

Lizenz- und Risikohinweise

Apache-2.0

ToolAI metadata-only listing: repository metadata is public; editorial content and manual verification are still pending.

sglang

sgl-project/sglang

★ 32,2KPython

OpenVINO

openvinotoolkit/openvino

★ 10,6KC++

GPUStack

gpustack/gpustack

★ 5,4KPython

FastDeploy

PaddlePaddle/FastDeploy

★ 3,7KPython