ToolAI.io · GitHub-kanaal

Multimodale AI · Open-source-AI-projecten op GitHub

Een feitengebaseerde index van LLM-, agent-, MCP-, RAG- en AI-ontwikkelingsprojecten, met licentie-, installatie-, download- en screenshotinformatie en gerelateerde bronnen voor elke vermelding.

Openbare repositorygegevens

Projectindex

14 Projecten
Multimodale AI

UniRL

Tencent-Hunyuan/UniRL

UniRL is a Framework for Unified Multimodal Model Reinforcement Learning

★ 854⑂ 56Python
NOASSERTIONQ94
Multimodale AI

MoneyPrinterTurbo

harry0703/MoneyPrinterTurbo

An AI-assisted workflow for generating short videos from a topic or keywords.

★ 120,7K⑂ 18,5KPython
MITQ90
Multimodale AI

GLM-V

zai-org/GLM-V

GLM-4.6V/4.5V/4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

★ 2,4K⑂ 174Python
Apache-2.0Q87
Schermafbeelding van PaddleOCR — OCR and document intelligence
Multimodale AI

PaddleOCR — OCR and document intelligence

PaddlePaddle/PaddleOCR

Turn any PDF or image document into structured data for your AI. A powerful, lightweight OCR toolkit that bridges the gap between images/PDFs and LLMs. Supports 100+ languages

★ 88K⑂ 11,2KPython
Apache-2.0Q86
Schermafbeelding van Qwen3-VL — Vision-language models
Multimodale AI

Qwen3-VL — Vision-language models

QwenLM/Qwen3-VL

Qwen3-VL is the multimodal large language model series developed by Qwen team, Alibaba Cloud

★ 19,8K⑂ 1,8KPython
Apache-2.0Q86
Multimodale AI

HY-World-2.0

Tencent-Hunyuan/HY-World-2.0

HY-World 2.0: A Multi-Modal World Model for Reconstructing, Generating, and Simulating 3D Worlds

★ 2,4K⑂ 209Python
NOASSERTIONQ86
Multimodale AI

stable-diffusion-webui

AUTOMATIC1111/stable-diffusion-webui

A browser interface for Stable Diffusion image generation and extensions.

★ 164,8K⑂ 30,6KPython
AGPL-3.0Q85
Multimodale AI

HunyuanWorld-1.0

Tencent-Hunyuan/HunyuanWorld-1.0

Generating Immersive, Explorable, and Interactive 3D Worlds from Words or Pixels with Hunyuan3D World Model

★ 2,9K⑂ 261Python
NOASSERTIONQ85
Multimodale AI

HunyuanImage-3.0

Tencent-Hunyuan/HunyuanImage-3.0

HunyuanImage-3.0: A Powerful Native Multimodal Model for Image Generation

★ 3,2K⑂ 175Python
NOASSERTIONQ84

Pagina 1 / 2 · 14 projecten

Recent bijgewerkt

cherry-studioCherryHQ/cherry-studio★ 51,5K siyuansiyuan-note/siyuan★ 46,2K career-opscareer-ops-hq/career-ops★ 70,2K tensorflowtensorflow/tensorflow★ 198,8K streamlitstreamlit/streamlit★ 45,7K pytorchpytorch/pytorch★ 102,8K

Meeste sterren

ECCaffaan-m/ECC★ 248,8K Hermes Agentnousresearch/hermes-agent★ 227,1K tensorflowtensorflow/tensorflow★ 198,8K AutoGPTSignificant-Gravitas/AutoGPT★ 187,1K ollamaollama/ollama★ 180,2K markitdown — Document conversion and extractionmicrosoft/markitdown★ 174,9K