Inference, Deployment & Runtime

LocalAI

mudler/LocalAI

A local inference engine for self-hosting models and AI services.

★ 48.9KStars
⑂ 4.4KForks
214Open issues
GoLanguage
MITLicense
Q85Editorial score

Overview

A local inference engine for self-hosting models and AI services. The repository is maintained under mudler on GitHub. Its primary language is Go. See the [project README](https://github.com/mudler/LocalAI#readme) for the supported workflows.

Key features

  • A local inference engine for self-hosting models and AI services.

Requirements, installation and quick start

Use the upstream download or installation instructions for your operating system. Check the model, memory and accelerator requirements for the workload you intend to run.

[Read the upstream installation and quickstart instructions](https://github.com/mudler/LocalAI#readme).

Usage

Start with the documented local example, confirm it runs with your resources, then connect it to your application using the supported interface.

[Usage examples and configuration reference](https://github.com/mudler/LocalAI#readme).

Model compatibility and use cases

Model compatibility is not stated in the repository metadata.

License and risk notes

GitHub reports MIT for this repository. Review the upstream license file; model weights, datasets and dependencies may have separate terms.

Source review 2026-09-05: GitHub search metadata and repository README. No runtime benchmark performed. License metadata: MIT.

Release and maintenance

[View upstream releases](https://github.com/mudler/LocalAI/releases).

vllm

vllm-project/vllm

★ 89.5KPython

sglang

sgl-project/sglang

★ 32.2KPython

OpenVINO

openvinotoolkit/openvino

★ 10.6KC++

GPUStack

gpustack/gpustack

★ 5.4KPython