ایجنٹس اور ملٹی ایجنٹ
cloudgeni-ai/opengeni
An open, self-hostable agentic runtime for organizations that provides durable, replayable agent sessions, human approvals, governed credentials and memory, and flexible compute targets including managed sandboxes and enrolled user hardware.
★ 56⑂ 5TypeScript
Apache-2.0Q86
ایجنٹس اور ملٹی ایجنٹ
reyamira/models
A TUI and CLI tool for browsing AI models, benchmarks, coding agents, and provider statuses.
★ 492⑂ 19Rust
MITQ92
انفرنس، ڈپلائمنٹ اور رَن ٹائم
xllm-ai/xllm
xLLM is a C++ inference framework for LLM, VLM, DiT and REC models. It targets high-throughput, low-latency deployment on several AI accelerator families and separates service-layer scheduling and availability from engine-layer computation.
★ 1.5K⑂ 279C++
Apache-2.0Q98
ایجنٹس اور ملٹی ایجنٹ
future-agi/future-agi
An open-source, self-hostable platform for evaluating, tracing, simulating, protecting, routing, and optimizing LLM and AI-agent applications.
★ 1.7K⑂ 498Python
Apache-2.0Q98
MCP اور ٹول کالنگ
artokun/comfyui-mcp
A local-first, agent-native control plane for ComfyUI that provides an MCP server and sidebar agent to generate images, video, and audio, author and run workflows, and edit live graphs using natural language across any LLM.
★ 599⑂ 93TypeScript
MITQ92
MCP اور ٹول کالنگ
juspay/neurolink
A TypeScript integration platform providing a unified API for 30+ AI providers and 100+ models, enabling provider swapping, multi-modal voice processing, RAG, memory, and MCP-native tool integration.
★ 121⑂ 124TypeScript
MITQ93
انفرنس، ڈپلائمنٹ اور رَن ٹائم
vllm-project/vllm-ascend
A community-maintained hardware plugin that enables vLLM to run model inference and serving workloads on supported Ascend NPUs.
★ 2.7K⑂ 2.1KC++
Apache-2.0Q98
ایجنٹس اور ملٹی ایجنٹ
langwatch/langwatch
An open-core platform for evaluating, testing, tracing, and monitoring LLM applications and AI agents before release and in production.
★ 3.5K⑂ 345TypeScript
Apache-2.0Q98
انفرنس، ڈپلائمنٹ اور رَن ٹائم
avifenesh/memra
A from-scratch LLM inference engine built in Rust and CUDA, specifically optimized for RTX 5090 (sm_120a) and H100 (sm_90a) architectures, delivering exactness-gated performance without relying on frameworks like ggml.
★ 314⑂ 35Rust
MITQ92
LLM اور فاؤنڈیشن ماڈلز
bionic-gpt/bionic-gpt
Bionic is a Rust-based, on-premise alternative to ChatGPT designed for private generative AI deployments. It provides a familiar chat interface, local or remote model access, team controls, retrieval-augmented assistants, data integrations, observability, and Kubernetes-oriented scaling.
★ 2.4K⑂ 239Rust
NOASSERTIONQ98
انفرنس، ڈپلائمنٹ اور رَن ٹائم
mozilla-ai/any-llm
any-llm is a Python SDK for communicating with multiple LLM providers through one interface. It supports switching providers and models with minimal code changes while using official provider SDKs when available.
★ 2.2K⑂ 211Python
Apache-2.0Q98
فائن ٹیوننگ، تربیت اور ڈیٹا
kaito-project/aikit
🏗️ Fine-tune, build, and deploy open-source LLMs easily!
★ 535⑂ 57Go
MITQ94