38 repos
Tools and models for converting text into spoken audio across multiple languages, with emphasis on neural TTS systems and integration with large language models. The cluster includes multilingual implementations (Mandarin, Portuguese, Hindi, Spanish, and English), foundational TTS architectures like T5Gemma-TTS, and related speech synthesis research. Projects here range from production-ready TTS engines to experimental model architectures and language-specific voice synthesis variants.