Inferencia, implementación y tiempo de ejecución
mozilla-ai/any-llm
any-llm is a Python SDK for communicating with multiple LLM providers through one interface. It supports switching providers and models with minimal code changes while using official provider SDKs when available.
★ 2,2K⑂ 211Python
Apache-2.0Q98
MCP y llamadas a herramientas
askimo-ai/askimo
A native desktop AI client for chat, local RAG, multi-step AI workflows (Plans), and agent skills, supporting multiple cloud and local LLM providers while keeping user files strictly on the machine.
★ 334⑂ 71Kotlin
AGPL-3.0Q94
Inferencia, implementación y tiempo de ejecución
xllm-ai/xllm
xLLM is a C++ inference framework for LLM, VLM, DiT and REC models. It targets high-throughput, low-latency deployment on several AI accelerator families and separates service-layer scheduling and availability from engine-layer computation.
★ 1,5K⑂ 279C++
Apache-2.0Q98
MCP y llamadas a herramientas
juspay/neurolink
A TypeScript integration platform providing a unified API for 30+ AI providers and 100+ models, enabling provider swapping, multi-modal voice processing, RAG, memory, and MCP-native tool integration.
★ 121⑂ 124TypeScript
MITQ93
Agentes y multiagente
0xlazai/alith
A simple, composable, and high-performance AI agent framework designed for Web3 and Crypto, enabling developers to build, deploy, and manage on-chain AI agents with multi-language support and LazAI Gateway integration.
★ 44⑂ 31Rust
Apache-2.0Q88
MCP y llamadas a herramientas
mtrnix/metronix-memory
Self-hosted memory infrastructure for AI agents featuring MCP-native integration, hybrid RAG, a temporal knowledge graph, and an ontology layer, designed for local-model friendliness and durable, agent-scoped context.
★ 39⑂ 7Python
Apache-2.0Q86
Agentes y multiagente
sno-ai/llmix
A production LLM call layer for AI agents and tools that wraps existing provider SDKs with config-driven model presets, caching, resilience patterns, and key rotation across Python, TypeScript, and Rust.
★ 131⑂ 28Python
Apache-2.0Q91
RAG y sistemas de conocimiento
stackitcloud/rag-template
A template for building AI chatbots and document management systems using Retrieval-Augmented Generation (RAG), vector search, and FastAPI, designed for deployment on Kubernetes.
★ 86⑂ 10Python
Apache-2.0Q89
Agentes y multiagente
microsoft/SkillOpt
SkillOpt is a text-space optimizer that trains reusable natural-language skills for frozen LLM agents through trajectory-driven edits, validation-gated updates, and deployable best_skill.md artifacts
★ 16,1K⑂ 1,5KPython
MITQ86
Evaluación, observabilidad y seguridad
NVIDIA/garak
the LLM vulnerability scanner
★ 8,8K⑂ 1,2KPython
Apache-2.0Q84
MCP y llamadas a herramientas
langchain4j/langchain4j
An idiomatic, open-source Java library for building LLM-powered applications on the JVM, offering a unified API over popular LLM providers and vector stores with support for tool calling, MCP, agents, and RAG.
★ 12,8K⑂ 2,4KJava
Apache-2.0Q98
Inferencia, implementación y tiempo de ejecución
xorbitsai/inference
A powerful and versatile library designed to serve language, speech recognition, and multimodal models. It allows users to swap GPT for any LLM by changing a single line of code and run models on cloud, on-prem, or locally via a unified, production-ready inference API.
★ 9,5K⑂ 853Python
Apache-2.0Q98