推理、部署与运行时

vllm

vllm-project/vllm

A high-throughput and memory-efficient inference and serving engine for LLMs

★ 89.4K星级
⑂ 20.9KFork 数
6746未解决的问题
Python语言
Apache-2.0许可证
Q@project.QualityScore编辑评分

概览

A high-throughput and memory-efficient inference and serving engine for LLMs

要求、安装和快速入门

安装要求未在仓库元数据中说明。

用法

用法未在仓库元数据中注明。

模型兼容性与使用场景

仓库元数据未说明模型兼容性。

许可证与风险说明

Apache-2.0

ToolAI metadata-only listing: repository metadata is public; editorial content and manual verification are still pending.

sglang

sgl-project/sglang

★ 32.1KPython

OpenVINO

openvinotoolkit/openvino

★ 10.6KC++

GPUStack

gpustack/gpustack

★ 5.4KPython