ملٹی موڈل AI
Qwen3-VL — Vision-language models
QwenLM/Qwen3-VL
Qwen3-VL is the multimodal large language model series developed by Qwen team, Alibaba Cloud
★ 19.8K⑂ 1.8KPython
ToolAI.io · GitHub چینل
GitHub پر حال ہی میں شامل کیے گئے اوپن سورس AI پروجیکٹس دریافت کریں، جن کے ساتھ ریپوزٹری کی معلومات، سیٹ اپ نوٹس، لائسنس اور متعلقہ وسائل بھی موجود ہیں۔
عوامی ریپوزٹری کا ڈیٹا
QwenLM/Qwen3-VL
Qwen3-VL is the multimodal large language model series developed by Qwen team, Alibaba Cloud
deepseek-ai/DeepSeek-OCR
Contexts Optical Compression
deepseek-ai/Janus
Janus-Series: Unified Multimodal Understanding and Generation Models
PaddlePaddle/PaddleOCR
Turn any PDF or image document into structured data for your AI. A powerful, lightweight OCR toolkit that bridges the gap between images/PDFs and LLMs. Supports 100+ languages
صفحہ 3 / 3 · 28 پروجیکٹس