ollama
ollama/ollama
A tool for running and managing large language models locally.
ToolAI.io · Chaîne GitHub
Un index factuel des projets de développement LLM, agent, MCP, RAG et AI, avec pour chaque entrée la licence, la configuration, le téléchargement, des captures d’écran et les ressources associées.
Données du dépôt public
ollama/ollama
A tool for running and managing large language models locally.
hal0ai/hal0
An open-source, self-hosted home AI inference platform designed to turn a Linux box into an OpenAI-compatible inference appliance, with native optimization for AMD Strix Halo hardware.
NVIDIA/Model-Optimizer
A unified library of SOTA model optimization techniques like quantization, distillation, pruning, neural architecture search, speculative decoding, etc. It compresses deep learning models for downstream deployment frameworks like TensorRT-LLM, TensorRT, vLLM, etc. to optimize inference speed.
mudler/LocalAI
A local inference engine for self-hosting models and AI services.
ray-project/ray
A distributed computing runtime for machine learning training, tuning and model serving.
MoonshotAI/checkpoint-engine
Checkpoint-engine is a simple middleware to update model weights in LLM inference engines
janhq/jan
A local AI chat application for running models on a personal computer.
Page 2 / 2 · 19 projets