इन्फरेंस, डिप्लॉयमेंट और रनटाइम
xorbitsai/inference
A powerful and versatile library designed to serve language, speech recognition, and multimodal models. It allows users to swap GPT for any LLM by changing a single line of code and run models on cloud, on-prem, or locally via a unified, production-ready inference API.
★ 9.5K⑂ 853Python
Apache-2.0Q98
इन्फरेंस, डिप्लॉयमेंट और रनटाइम
hal0ai/hal0
An open-source, self-hosted home AI inference platform designed to turn a Linux box into an OpenAI-compatible inference appliance, with native optimization for AMD Strix Halo hardware.
★ 68⑂ 7Python
Apache-2.0Q89
इन्फरेंस, डिप्लॉयमेंट और रनटाइम
janhq/jan
A local AI chat application for running models on a personal computer.
★ 44.3K⑂ 3KTypeScript
लाइसेंस का पता नहीं चलाQ80