ToolAI.io · Canal de GitHub

Proyectos de IA en tendencia en GitHub

Sigue repositorios de AI de código abierto activos en los ámbitos de LLM, agentes, MCP, RAG y programación, con su número de estrellas y datos de proyectos verificados.

Datos del repositorio público

Índice del proyecto

130 Proyectos
Captura de pantalla de memra
Inferencia, implementación y tiempo de ejecución

memra

avifenesh/memra

A from-scratch LLM inference engine built in Rust and CUDA, specifically optimized for RTX 5090 (sm_120a) and H100 (sm_90a) architectures, delivering exactness-gated performance without relying on frameworks like ggml.

★ 312⑂ 35Rust
MITQ92
Captura de pantalla de 1flowbase
Agentes y multiagente

1flowbase

taichuy/1flowbase

An open-source AI gateway that allows local agent clients to publish fusion-style multi-model workflows as OpenAI and Claude-compatible virtual models, providing full observability into traces, tokens, latency, and costs.

★ 257⑂ 19Rust
Apache-2.0Q89
Captura de pantalla de SBproxy
MCP y llamadas a herramientas

SBproxy

soapbucket/sbproxy

An open-source, self-hosted Enterprise AI Gateway and LLM proxy written in Rust, providing an OpenAI-compatible API for over 60 providers and supporting local model hosting, MCP tool federation, and robust traffic governance.

★ 49⑂ 1Rust
Apache-2.0Q86
Captura de pantalla de TensorRT-LLM — Optimized LLM inference
Inferencia, implementación y tiempo de ejecución

TensorRT-LLM — Optimized LLM inference

NVIDIA/TensorRT-LLM

TensorRT LLM provides users with an easy-to-use Python API to define Large Language Models (LLMs) and supports state-of-the-art optimizations to perform inference efficiently on NVIDIA GPUs. TensorRT LLM also contains components to create Python and C++ runtimes that orchestrate the inference execution in a performant way

★ 14,4K⑂ 2,7KC++
NOASSERTIONQ82
Agentes y multiagente

dstack

dstackai/dstack

A vendor-agnostic control plane for provisioning and orchestrating training, inference, development, and agentic workloads across GPU clouds, Kubernetes, and on-premises infrastructure.

★ 2,2K⑂ 248Python
MPL-2.0Q98
LLM y modelos fundacionales

Bionic GPT

bionic-gpt/bionic-gpt

Bionic is a Rust-based, on-premise alternative to ChatGPT designed for private generative AI deployments. It provides a familiar chat interface, local or remote model access, team controls, retrieval-augmented assistants, data integrations, observability, and Kubernetes-oriented scaling.

★ 2,4K⑂ 239Rust
NOASSERTIONQ98
Captura de pantalla de DotCraft
Agentes y multiagente

DotCraft

dotharness/dotcraft

An open-source, self-hosted, project-scoped AI agent runtime that provides persistent sessions, memory, background work, and automations shared across Desktop, CLI, bots, and applications.

★ 385⑂ 16C#
Apache-2.0Q88
Inferencia, implementación y tiempo de ejecución

vLLM Ascend

vllm-project/vllm-ascend

A community-maintained hardware plugin that enables vLLM to run model inference and serving workloads on supported Ascend NPUs.

★ 2,7K⑂ 2,1KC++
Apache-2.0Q98
Captura de pantalla de comfyui-mcp
MCP y llamadas a herramientas

comfyui-mcp

artokun/comfyui-mcp

A local-first, agent-native control plane for ComfyUI that provides an MCP server and sidebar agent to generate images, video, and audio, author and run workflows, and edit live graphs using natural language across any LLM.

★ 597⑂ 93TypeScript
MITQ92
Captura de pantalla de Infino
RAG y sistemas de conocimiento

Infino

infino-ai/infino

Infino is a fast retrieval engine that executes SQL, full-text search, and vector search over a single copy of data stored natively as Parquet on object storage.

★ 67⑂ 18Rust
Apache-2.0Q89

Página 2 / 11 · 130 proyectos

Actualizado recientemente

codex — Coding agent and developer workflowsopenai/codex★ 106,5K NemoClaw — Agent runtime and deployment toolingNVIDIA/NemoClaw★ 22,2K XERJxerj-org/xerj★ 1,4K LangWatchlangwatch/langwatch★ 3,5K OrchestKityonatangross/orchestkit★ 222 hal0hal0ai/hal0★ 67

Con más estrellas

Hermes Agentnousresearch/hermes-agent★ 227,1K markitdown — Document conversion and extractionmicrosoft/markitdown★ 174,2K skills — Reusable agent skills and workflowsanthropics/skills★ 169,9K Hugging Face Transformershuggingface/transformers★ 163,3K Firecrawlfirecrawl/firecrawl★ 161,1K LangChainlangchain-ai/langchain★ 143,6K