Qwen3-TTS: multilingual speech, voice design and cloning
QwenLM/Qwen3-TTS
An open speech model family from Qwen with preset speakers, text-driven voice design and reference-audio cloning for narration and speech applications.
ToolAI.io · Chaîne GitHub
Découvrez les nouveaux projets AI open source ajoutés sur GitHub, avec les informations du dépôt, les instructions de configuration, les licences et les ressources associées.
Données du dépôt public
QwenLM/Qwen3-TTS
An open speech model family from Qwen with preset speakers, text-driven voice design and reference-audio cloning for narration and speech applications.
PaddlePaddle/PaddleSpeech
A PaddlePaddle-based speech toolkit covering speech recognition, text-to-speech, punctuation restoration, and other audio tasks, with command-line, Python interface, and service examples.
facebookresearch/audiocraft
A PyTorch-based audio generation research library containing components such as MusicGen, AudioGen, and EnCodec. It supports text-conditioned generation, some melody-conditioned tasks, and model training workflows.
2noise/ChatTTS
Conversational Chinese and English speech generation with speaker and prosody controls; released model weights are noncommercial.
CorentinJ/Real-Time-Voice-Cloning
A speech synthesis project for exploring speaker representations and voice cloning.
openai/openai-agents-js
A lightweight, powerful framework for multi-agent workflows and voice agents
juspay/neurolink
A TypeScript integration platform providing a unified API for 30+ AI providers and 100+ models, enabling provider swapping, multi-modal voice processing, RAG, memory, and MCP-native tool integration.