Text-to-Speech and Voice Synthesis

6 repos

Tools, models, and frameworks for converting written text into natural-sounding speech across multiple languages. The cluster centers on transformer-based TTS architectures, model quantization formats like GGUF, and multilingual implementations including Chinese, Portuguese, Hindi, and other language variants. Repositories range from inference engines and model weights to complete end-to-end TTS systems, with emphasis on open-source, efficient implementations suitable for both research and production deployment.

text-to-speech ·198
transformers ·159
safetensors ·159
tts ·159
speech ·124
t5gemma_voice ·124
custom_code ·124
text2text-generation ·124
male ·39
female ·39