Text-to-Speech and Voice Synthesis

9 repos

Tools and models for converting text into natural-sounding speech, including zero-shot TTS systems that can synthesize speech from minimal data. The cluster centers on open-source TTS implementations, benchmarking frameworks for evaluating voice synthesis quality, and model evaluation pipelines. Repositories here focus on building, training, and assessing neural speech synthesis models with an emphasis on accessibility and reproducibility.

Python · 1
text-to-speech ·267
tts ·240
speech-synthesis ·217
zero-shot-tts ·171
english ·156
voice-cloning ·106
flow-matching ·106
diffusion-transformer ·106
en ·106
f5-tts ·106