Qwen3-TTS: multilingual speech, voice design and cloning
QwenLM/Qwen3-TTS
An open speech model family from Qwen with preset speakers, text-driven voice design and reference-audio cloning for narration and speech applications.
ToolAI.io · Kanal GitHub
Temukan proyek AI sumber terbuka yang baru ditambahkan di GitHub, lengkap dengan informasi repositori, catatan penyiapan, lisensi, dan sumber daya terkait.
Data repositori publik
QwenLM/Qwen3-TTS
An open speech model family from Qwen with preset speakers, text-driven voice design and reference-audio cloning for narration and speech applications.
PaddlePaddle/PaddleSpeech
A PaddlePaddle-based speech toolkit covering speech recognition, text-to-speech, punctuation restoration, and other audio tasks, with command-line, Python interface, and service examples.
google-deepmind/mujoco
A multijoint contact dynamics engine maintained by Google DeepMind, providing model description, physics stepping, an interactive viewer, and a Python interface. It is suitable for developing robotics, control, and reinforcement learning environments.
hydra-ecosystem/hydra
A configuration framework for Python applications that manages experiments through configuration composition, command-line overrides, and multirun execution. It is suitable for organizing data, model, and training parameters into reusable configuration structures.
facebookresearch/vggt
Visual Geometry Grounded Transformer predicts camera parameters, depth maps, point maps, and point tracks from scene images, making it suitable for preliminary geometric estimation in 3D vision research and reconstruction workflows.
facebookresearch/audiocraft
A PyTorch-based audio generation research library containing components such as MusicGen, AudioGen, and EnCodec. It supports text-conditioned generation, some melody-conditioned tasks, and model training workflows.
facebookresearch/segment-anything
Meta’s SAM image segmentation project supports prompt inputs such as points and boxes, and can also automatically generate candidate masks for an entire image. It is suitable for interactive annotation and image processing workflows.
anthropics/claude-cookbooks
Anthropic’s collection of Claude development examples covers classification, summarization, retrieval augmentation, tool use, image understanding, and evaluation. It is suitable for learning from specific tasks and adapting the code.
microsoft/playwright
Use a single API to drive Chromium, Firefox, and WebKit for web testing, form interactions, screenshots, and failure replay; it can also serve as the execution foundation for AI browser tasks.
Lightning-AI/pytorch-lightning
Organize PyTorch training with reusable loops, checkpoints, logging, and device strategies.
hpcaitech/ColossalAI
Distributed training and inference tools for parallelism and memory management in large-model workloads.
huggingface/diffusers
A Python library for pretrained diffusion pipelines, reusable components, and training examples.
Halaman 2 / 25 · 295 proyek