9 repos
Tools and models for converting text into natural-sounding speech, including zero-shot TTS systems that can synthesize speech from minimal data. The cluster centers on open-source TTS implementations, benchmarking frameworks for evaluating voice synthesis quality, and model evaluation pipelines. Repositories here focus on building, training, and assessing neural speech synthesis models with an emphasis on accessibility and reproducibility.