vllm
vllm-project/vllm
A high-throughput and memory-efficient inference and serving engine for LLMs
ToolAI.io · Канал GitHub
Фактический каталог проектов по разработке LLM, агентов, MCP, RAG и AI с информацией о лицензии, настройке, скачивании, скриншотами и связанными ресурсами для каждой записи.
Python tool for converting files and office documents to Markdown
★ 174,9K · PythonPublic repository for Agent Skills
★ 170,6K · MarkdownClaude Code is an agentic coding tool that lives in your terminal, understands your codebase, and helps you code faster by executing routine tasks, explaining complex code, and handling git workflows - all through natural language commands
★ 142,1K · TypeScriptRobust Speech Recognition via Large-Scale Weak Supervision
★ 107,7K · PythonAn open-source project focused on large language model research and inference.
★ 104,4K · PythonAn open-source project focused on reasoning model research and evaluation.
★ 92K · PythonA self-improving AI agent built by Nous Research that features a closed learning loop, multi-platform messaging integration, and flexible model provider switching. It is designed to run on low-cost VPS, GPU clusters, or serverless infrastructure.
★ 227,1K · PythonA model-definition framework for state-of-the-art machine learning models across text, vision, audio, and multimodal domains, supporting both inference and training.
★ 163,3K · PythonAn open-source web context API to search, scrape, and interact with the web at scale, converting web data into clean Markdown or structured JSON for AI agents and LLMs.
★ 161,1K · TypeScript MCP и вызов инструментовDify is an LLM application development platform for visually building workflows, RAG pipelines, and tool-using agents. It combines model management, prompt development, observability, and application APIs in a collaborative workspace that can be used as a hosted service or self-hosted.
★ 151,2K · TypeScriptAn open-source Python framework for building LLM-powered applications and agents through interoperable, chainable components and third-party integrations.
★ 143,6K · PythonGraphify is a local-first CLI tool and AI assistant skill that parses codebases, docs, SQL schemas, and PDFs into a queryable knowledge graph using deterministic AST parsing, eliminating the need for vector stores or embeddings.
★ 104K · PythonДанные публичного репозитория
vllm-project/vllm
A high-throughput and memory-efficient inference and serving engine for LLMs
infiniflow/ragflow
RAGFlow is an open-source Retrieval-Augmented Generation engine that combines document ingestion, retrieval, grounded citations, configurable language and embedding models, and agent capabilities to provide context for LLM applications.
openhands/openhands
A self-hosted developer control center for running, managing, and automating coding agents across local, remote, and cloud environments.
unslothai/unsloth
Unsloth is a Python-based, self-hosted toolkit with a beta web UI and code-based core for running, fine-tuning, exporting, and serving language, vision, audio, and embedding models locally.
headroomlabs-ai/headroom
Headroom is a local-first context compression layer that reduces token usage for AI agents by compressing tool outputs, logs, files, and RAG chunks before they reach the LLM. It offers library, proxy, and MCP server modes and claims to preserve answer accuracy while cutting tokens by 15-95% depending on workload.
mintplex-labs/anything-llm
An all-in-one, local-first AI application for chatting with documents, building AI agents, and running a private, multi-user ChatGPT-like experience with zero setup friction.
mem0ai/mem0
Mem0 is an open-source memory layer for AI agents and assistants, enabling personalized, long-term memory across user sessions. It provides multi-level memory retention, hybrid retrieval, and temporal reasoning through Python and npm SDKs, a self-hosted server, or a managed cloud platform.
mempalace/mempalace
A local-first, open-source AI memory system that stores conversation history and project files as verbatim text for semantic retrieval, achieving 96.6% R@5 on LongMemEval without requiring LLMs or API calls.
berriai/litellm
The fastest, litest AI Gateway. Rust core with Python SDK. Call 100+ LLM APIs in OpenAI (or native) format with cost tracking, guardrails, load balancing, and logging [Bedrock, Azure, OpenAI, Anthropic, OpenAI, VertexAI, vLLM, Nvidia NIM]
run-llama/llama_index
LlamaIndex is an open-source Python data framework for building LLM applications with private data, offering data connectors, indexing structures, and advanced retrieval interfaces. It is complemented by LlamaParse, an enterprise platform for agentic OCR, parsing, extraction, and indexing.
milvus-io/milvus
Milvus is a high-performance, cloud-native vector database built for scalable vector ANN search
astrbotdevs/astrbot
AstrBot is an open-source, all-in-one AI agent chatbot platform and development framework that integrates mainstream instant messaging platforms, large language models, plugins, and agent capabilities for building production-ready conversational AI applications.
Страница 2 / 22 · 253 проектов