इन्फरेंस, डिप्लॉयमेंट और रनटाइम

vllm

vllm-project/vllm

A high-throughput and memory-efficient inference and serving engine for LLMs

★ 89.5Kसितारे
⑂ 21Kफ़ोर्क्स
6832खुले मुद्दे
Pythonभाषा
Apache-2.0लाइसेंस
Q98संपादकीय स्कोर

अवलोकन

A high-throughput and memory-efficient inference and serving engine for LLMs

आवश्यकताएँ, इंस्टॉलेशन और त्वरित शुरुआत

रिपॉज़िटरी मेटाडेटा में इंस्टॉलेशन आवश्यकताओं का उल्लेख नहीं है।

उपयोग

रिपॉज़िटरी मेटाडेटा में उपयोग का उल्लेख नहीं है।

मॉडल संगतता और उपयोग के मामले

रिपॉज़िटरी मेटाडेटा में मॉडल संगतता का उल्लेख नहीं है।

लाइसेंस और जोखिम संबंधी टिप्पणियाँ

Apache-2.0

ToolAI metadata-only listing: repository metadata is public; editorial content and manual verification are still pending.

sglang

sgl-project/sglang

★ 32.2KPython

OpenVINO

openvinotoolkit/openvino

★ 10.6KC++

GPUStack

gpustack/gpustack

★ 5.4KPython

FastDeploy

PaddlePaddle/FastDeploy

★ 3.7KPython