Isaac GR00T: vision-language-action models for robotics
NVIDIA/Isaac-GR00T
NVIDIA’s robot models and reference code offer multimodal action prediction, demonstration-data adaptation, fine-tuning and deployment tooling.
ToolAI.io · Kênh GitHub
Khám phá các dự án AI nguồn mở mới được thêm trên GitHub, cùng thông tin kho lưu trữ, hướng dẫn thiết lập, giấy phép và tài nguyên liên quan.
Dữ liệu kho lưu trữ công khai
NVIDIA/Isaac-GR00T
NVIDIA’s robot models and reference code offer multimodal action prediction, demonstration-data adaptation, fine-tuning and deployment tooling.
microsoft/TRELLIS.2
Microsoft’s 3D generation project with image-to-3D inference, PBR texturing and GLB export for GPU-equipped asset workflows.
QwenLM/Qwen-Image
Qwen image generation and editing models with Diffusers examples for text-rich artwork, product imagery and reference-guided edits.
huggingface/diffusers
A Python library for pretrained diffusion pipelines, reusable components, and training examples.
harry0703/MoneyPrinterTurbo
An AI-assisted workflow for generating short videos from a topic or keywords.
AUTOMATIC1111/stable-diffusion-webui
A browser interface for Stable Diffusion image generation and extensions.
zai-org/CogVideo
text and image to video generation: CogVideoX (2024) and CogVideo (ICLR 2023)
Tencent-Hunyuan/Hunyuan3D-2
High-Resolution 3D Assets Generation with Large Scale Hunyuan3D Diffusion Models.
Tencent-Hunyuan/HY-World-2.0
HY-World 2.0: A Multi-Modal World Model for Reconstructing, Generating, and Simulating 3D Worlds
Tencent-Hunyuan/HunyuanWorld-1.0
Generating Immersive, Explorable, and Interactive 3D Worlds from Words or Pixels with Hunyuan3D World Model
zai-org/GLM-V
GLM-4.6V/4.5V/4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning
Tencent-Hunyuan/HunyuanImage-3.0
HunyuanImage-3.0: A Powerful Native Multimodal Model for Image Generation
Trang 1 / 3 · 25 dự án