추론, 배포 및 런타임
mozilla-ai/any-llm
any-llm is a Python SDK for communicating with multiple LLM providers through one interface. It supports switching providers and models with minimal code changes while using official provider SDKs when available.
★ 2.2K⑂ 211Python
Apache-2.0Q98
MCP 및 도구 호출
askimo-ai/askimo
A native desktop AI client for chat, local RAG, multi-step AI workflows (Plans), and agent skills, supporting multiple cloud and local LLM providers while keeping user files strictly on the machine.
★ 334⑂ 71Kotlin
AGPL-3.0Q94
추론, 배포 및 런타임
xllm-ai/xllm
xLLM is a C++ inference framework for LLM, VLM, DiT and REC models. It targets high-throughput, low-latency deployment on several AI accelerator families and separates service-layer scheduling and availability from engine-layer computation.
★ 1.5K⑂ 279C++
Apache-2.0Q98
MCP 및 도구 호출
juspay/neurolink
A TypeScript integration platform providing a unified API for 30+ AI providers and 100+ models, enabling provider swapping, multi-modal voice processing, RAG, memory, and MCP-native tool integration.
★ 121⑂ 124TypeScript
MITQ93
에이전트 및 멀티 에이전트
0xlazai/alith
A simple, composable, and high-performance AI agent framework designed for Web3 and Crypto, enabling developers to build, deploy, and manage on-chain AI agents with multi-language support and LazAI Gateway integration.
★ 44⑂ 31Rust
Apache-2.0Q88
MCP 및 도구 호출
mtrnix/metronix-memory
Self-hosted memory infrastructure for AI agents featuring MCP-native integration, hybrid RAG, a temporal knowledge graph, and an ontology layer, designed for local-model friendliness and durable, agent-scoped context.
★ 39⑂ 7Python
Apache-2.0Q86
에이전트 및 멀티 에이전트
sno-ai/llmix
A production LLM call layer for AI agents and tools that wraps existing provider SDKs with config-driven model presets, caching, resilience patterns, and key rotation across Python, TypeScript, and Rust.
★ 131⑂ 28Python
Apache-2.0Q91
RAG 및 지식 시스템
stackitcloud/rag-template
A template for building AI chatbots and document management systems using Retrieval-Augmented Generation (RAG), vector search, and FastAPI, designed for deployment on Kubernetes.
★ 86⑂ 10Python
Apache-2.0Q89
에이전트 및 멀티 에이전트
microsoft/SkillOpt
SkillOpt is a text-space optimizer that trains reusable natural-language skills for frozen LLM agents through trajectory-driven edits, validation-gated updates, and deployable best_skill.md artifacts
★ 16.1K⑂ 1.5KPython
MITQ86
평가, 옵저버빌리티 및 안전
NVIDIA/garak
the LLM vulnerability scanner
★ 8.8K⑂ 1.2KPython
Apache-2.0Q84
MCP 및 도구 호출
langchain4j/langchain4j
An idiomatic, open-source Java library for building LLM-powered applications on the JVM, offering a unified API over popular LLM providers and vector stores with support for tool calling, MCP, agents, and RAG.
★ 12.8K⑂ 2.4KJava
Apache-2.0Q98
추론, 배포 및 런타임
xorbitsai/inference
A powerful and versatile library designed to serve language, speech recognition, and multimodal models. It allows users to swap GPT for any LLM by changing a single line of code and run models on cloud, on-prem, or locally via a unified, production-ready inference API.
★ 9.5K⑂ 853Python
Apache-2.0Q98