推論、デプロイ、ランタイム

vllm

vllm-project/vllm

A high-throughput and memory-efficient inference and serving engine for LLMs

★ 89.4Kスター
⑂ 20.9Kフォーク数
6746未解決の問題
Python言語
Apache-2.0ライセンス
Q@project.QualityScore編集部スコア

概要

A high-throughput and memory-efficient inference and serving engine for LLMs

要件、インストール、クイックスタート

リポジトリのメタデータにインストール要件が記載されていません。

使用方法

使用方法はリポジトリのメタデータに記載されていません。

モデルの互換性とユースケース

リポジトリのメタデータにモデル互換性の記載がありません。

ライセンスとリスクに関する注意事項

Apache-2.0

ToolAI metadata-only listing: repository metadata is public; editorial content and manual verification are still pending.

sglang

sgl-project/sglang

★ 32.1KPython

OpenVINO

openvinotoolkit/openvino

★ 10.6KC++

GPUStack

gpustack/gpustack

★ 5.4KPython