AI đa phương thức
Qwen3-VL — Vision-language models
QwenLM/Qwen3-VL
Qwen3-VL is the multimodal large language model series developed by Qwen team, Alibaba Cloud
★ 19,8K⑂ 1,8KPython
ToolAI.io · Kênh GitHub
Khám phá các dự án AI nguồn mở mới được thêm trên GitHub, cùng thông tin kho lưu trữ, hướng dẫn thiết lập, giấy phép và tài nguyên liên quan.
Dữ liệu kho lưu trữ công khai
QwenLM/Qwen3-VL
Qwen3-VL is the multimodal large language model series developed by Qwen team, Alibaba Cloud
deepseek-ai/DeepSeek-OCR
Contexts Optical Compression
deepseek-ai/Janus
Janus-Series: Unified Multimodal Understanding and Generation Models
PaddlePaddle/PaddleOCR
Turn any PDF or image document into structured data for your AI. A powerful, lightweight OCR toolkit that bridges the gap between images/PDFs and LLMs. Supports 100+ languages
Trang 3 / 3 · 28 dự án