33 repos
Aratako/T5Gemma-TTS
Multilingual TTS model with voice cloning and duration control, based on T5Gemma encoder-decoder LLM
312
22 commits
Aratako/Irodori-TTS
A Flow Matching-based Text-to-Speech Model with Emoji-driven Style Control
1,297
35 commits
dots-studio/dots.tts-base
dots.tts-base
52
9 commits
rednote-hilab/dots.tts-base
maum-ai/univnet
Unofficial PyTorch Implementation of UnivNet Vocoder (https://arxiv.org/abs/2106.07889)
286
11 commits
franciscocarloserra/ttslibre
Small, fast, free: an open recipe for a sub-100M TTS. Code, data, weights.
38
38 commits
dots-studio/dots.tts-soar
dots.tts-soar
92
8 commits
rednote-hilab/dots.tts-soar
jik876/hifi-gan
HiFi-GAN: Generative Adversarial Networks for Efficient and High Fidelity Speech Synthesis
2,368
mozi1924/Qwen3-TTS-EasyFinetuning
Easy fine-tuning for Qwen3-TTS: Fast voice cloning and high-quality multilingual speech synthesis.
122
71 commits
rednote-hilab/dots.tts-mf
dots.tts-mf
28
7 commits
dots-studio/dots.tts-mf
coqui-ai/TTS
🐸💬 - a deep learning toolkit for Text-to-Speech, battle-tested in research and production
46,023
3,494 commits
HoseinAzad/SpeechT5-Non-English-TTS
Fine-tune SpeechT5 for non-English text-to-speech task, implemented in PyTorch.
8
14 commits
asiff00/Training-TTS
Train and finutune text-to-speech models for Bengali and many other languages!
18
gokhaneraslan/chatterbox-finetuning
Fine-tuning toolkit for Chatterbox TTS & Chatterbox TURBO models. Supports 23 languages with smart…
109
babysor/MockingBird
🚀Clone a voice in 5 seconds to generate arbitrary speech in real-time
36,908
206 commits
speechbrain/tts-hifigan-libritts-16kHz
Vocoder with HiFIGAN trained on LibriTTS
16 commits
speechbrain/tts-hifigan-ljspeech
Vocoder with HiFIGAN trained on LJSpeech
39
shivammehta25/Matcha-TTS
[ICASSP 2024] 🍵 Matcha-TTS: A fast TTS architecture with conditional flow matching
1,357
116 commits
lucasnewman/nanospeech
A simple, hackable text-to-speech system in PyTorch and MLX
190
23 commits
kadirnar/voicehub
VoiceHub: A Unified Inference Interface for TTS Models
119
260 commits
Camb-ai/mars5-tts
MARS5: A novel speech model for insane prosody.
480
24 commits
Chanson-0803/MSpoofTTS
MSpoofTTS Discriminator Checkpoints
1
6 commits
e-c-k-e-r/vall-e
An unofficial PyTorch implementation of VALL-E
88
877 commits
espnet/espnet
End-to-End Speech Processing Toolkit
9,964
24,634 commits
AutoArk/GPA
[AutoArk] GPA (General Purpose Audio) can do ASR, TTS and voice conversion with one tiny model!
3,107
gruporaia/TTS-AutoTuning
Pipeline para finetuning automático de modelos de Text to Speech.
0
86 commits
lucadellalib/audiocodecs
A collections of audio codecs with a standardized API
44
microsoft/speecht5_tts
SpeechT5 (TTS task)
845
26 commits
espnet/kan-bayashi_ljspeech_vits
ESPnet2 TTS pretrained model
226
2 commits
moonshotai/Kimi-Audio-7B-Instruct
Kimi-Audio
418
19 commits
CorentinJ/Real-Time-Voice-Cloning
Clone a voice in 5 seconds to generate arbitrary speech in real-time
60,142
274 commits