π§ ICE-012-Audio β multilingual TTS & voice cloning
5
3 commits
1 linked in READMEs
updated Sep 2, 2026
A Gradio demo for darkps/ice-012-audio,
a masked-diffusion (non-autoregressive) text-to-speech model with a Qwen3 backbone,
an acoustic prosody adapter, and the Higgs-Audio-v2 24 kHz neural codec.
What you can do
Runs on ZeroGPU.
examples/ljspeech_female.wav β LJSpeech (public domain).examples/librispeech_male.flac β LibriSpeech dev-clean utterance
1272-128104-0003, from LibriSpeech (CC-BY-4.0).Voice cloning should only be used with audio you own or have explicit permission to process. Do not use this demo to impersonate real people.
3 commits
π§ ICE-012-Audio β multilingual TTS & voice cloning
5
3 commits
1 linked in READMEs
updated Sep 2, 2026
A Gradio demo for darkps/ice-012-audio,
a masked-diffusion (non-autoregressive) text-to-speech model with a Qwen3 backbone,
an acoustic prosody adapter, and the Higgs-Audio-v2 24 kHz neural codec.
What you can do
Runs on ZeroGPU.
examples/ljspeech_female.wav β LJSpeech (public domain).examples/librispeech_male.flac β LibriSpeech dev-clean utterance
1272-128104-0003, from LibriSpeech (CC-BY-4.0).Voice cloning should only be used with audio you own or have explicit permission to process. Do not use this demo to impersonate real people.
3 commits