Text-to-Speech and Voice Generation

33 repos

Neural text-to-speech (TTS) systems and voice synthesis models that convert written text into spoken audio. The cluster centers on efficient TTS implementations including fish-speech, Qwen3-TTS, MOSS-TTS, and Bark—ranging from lightweight mobile-friendly models to customizable voice generation systems. These repositories span multilingual support (English, Italian, French, Korean, Spanish) and cover both pre-trained models and frameworks for building voice synthesis applications.

en ·13,702
text-to-speech ·13,702
de ·13,569
zh ·13,266
ko ·13,178
es ·12,230
it ·12,171
fr ·12,166
pt ·11,519
ru ·11,295

suno/bark

Model

Bark

1,564

25 commits