Speech, Audio & Voice
facebookresearch/audiocraft
A PyTorch-based audio generation research library containing components such as MusicGen, AudioGen, and EnCodec. It supports text-conditioned generation, some melody-conditioned tasks, and model training workflows.
★ 23.6K⑂ 2.7KPython
MITQ92
Speech, Audio & Voice
PaddlePaddle/PaddleSpeech
A PaddlePaddle-based speech toolkit covering speech recognition, text-to-speech, punctuation restoration, and other audio tasks, with command-line, Python interface, and service examples.
★ 12.7K⑂ 2KPython
Apache-2.0Q92
Speech, Audio & Voice
2noise/ChatTTS
Conversational Chinese and English speech generation with speaker and prosody controls; released model weights are noncommercial.
★ 39.8K⑂ 4.2KPython
AGPL-3.0Q90
Speech, Audio & Voice
CorentinJ/Real-Time-Voice-Cloning
A speech synthesis project for exploring speaker representations and voice cloning.
★ 60.1K⑂ 9.4KPython
License not detectedQ85
Speech, Audio & Voice
QwenLM/Qwen3-TTS
An open speech model family from Qwen with preset speakers, text-driven voice design and reference-audio cloning for narration and speech applications.
★ 13.4K⑂ 1.7KPython
Apache-2.0Q84