ToolAI.io · Canal de GitHub

AI multimodal · Proyectos de AI de código abierto en GitHub

Un índice basado en hechos de proyectos de desarrollo de LLM, agentes, MCP, RAG y AI, con licencia, configuración, descarga, capturas de pantalla y recursos relacionados para cada entrada.

Datos del repositorio público

Índice del proyecto

5 Proyectos
Captura de pantalla de PaddleOCR — OCR and document intelligence
AI multimodal

PaddleOCR — OCR and document intelligence

PaddlePaddle/PaddleOCR

Turn any PDF or image document into structured data for your AI. A powerful, lightweight OCR toolkit that bridges the gap between images/PDFs and LLMs. Supports 100+ languages

★ 87,8K⑂ 11,2KPython
Apache-2.0Q86
Captura de pantalla de Qwen3-VL — Vision-language models
AI multimodal

Qwen3-VL — Vision-language models

QwenLM/Qwen3-VL

Qwen3-VL is the multimodal large language model series developed by Qwen team, Alibaba Cloud

★ 19,8K⑂ 1,8KPython
Apache-2.0Q86

Actualizado recientemente

codex — Coding agent and developer workflowsopenai/codex★ 106,5K NemoClaw — Agent runtime and deployment toolingNVIDIA/NemoClaw★ 22,2K XERJxerj-org/xerj★ 1,4K LangWatchlangwatch/langwatch★ 3,5K OrchestKityonatangross/orchestkit★ 222 hal0hal0ai/hal0★ 67

Con más estrellas

Hermes Agentnousresearch/hermes-agent★ 227,1K markitdown — Document conversion and extractionmicrosoft/markitdown★ 174,2K skills — Reusable agent skills and workflowsanthropics/skills★ 169,9K Hugging Face Transformershuggingface/transformers★ 163,3K Firecrawlfirecrawl/firecrawl★ 161,1K LangChainlangchain-ai/langchain★ 143,6K