🎙️ Live demo: Try this model in the
ResembleAI/Chatterbox-Multilingual-TTS-es-esSpace.
Chatterbox Multilingual: Spanish (Spain) is a dedicated single-language finetune in the Chatterbox Multilingual V3 Single Language Pack. It is optimized for Spanish as spoken in Spain, with language- and region-specific behavior for expressive text-to-speech and voice cloning.
Use this model when you want tighter Spanish (Spain) quality control than the broad multilingual checkpoint. For a single model that covers all supported languages, use ResembleAI/chatterbox.
Try the hosted demo Space: ResembleAI/Chatterbox-Multilingual-TTS-es-es.
t3_es_es.safetensors: T3 state dict in safetensors format.s3gen_v3.pt / s3gen_v3.safetensors: V3 S3Gen speech decoder checkpoint.grapheme_mtl_merged_expanded_v1.json: multilingual tokenizer config.es-ESes135500t3_135500.pth.tar292float32(2454, 1024)(8194, 1024)2143990264 bytesd85844b13ea8cb45e95b8d84a55bcfeccb2d743035cf304ee3d778fc6be39546This repository contains Chatterbox Multilingual V3 single-language assets used by the linked demo Space. The T3 checkpoint is loaded with multilingual vocabulary shape 2454 and S3 speech vocabulary shape 8194.
The demo combines these model-specific assets with the shared Chatterbox inference code and companion assets needed for end-to-end speech generation.
🎙️ Live demo: Try this model in the
ResembleAI/Chatterbox-Multilingual-TTS-es-esSpace.
Chatterbox Multilingual: Spanish (Spain) is a dedicated single-language finetune in the Chatterbox Multilingual V3 Single Language Pack. It is optimized for Spanish as spoken in Spain, with language- and region-specific behavior for expressive text-to-speech and voice cloning.
Use this model when you want tighter Spanish (Spain) quality control than the broad multilingual checkpoint. For a single model that covers all supported languages, use ResembleAI/chatterbox.
Try the hosted demo Space: ResembleAI/Chatterbox-Multilingual-TTS-es-es.
t3_es_es.safetensors: T3 state dict in safetensors format.s3gen_v3.pt / s3gen_v3.safetensors: V3 S3Gen speech decoder checkpoint.grapheme_mtl_merged_expanded_v1.json: multilingual tokenizer config.es-ESes135500t3_135500.pth.tar292float32(2454, 1024)(8194, 1024)2143990264 bytesd85844b13ea8cb45e95b8d84a55bcfeccb2d743035cf304ee3d778fc6be39546This repository contains Chatterbox Multilingual V3 single-language assets used by the linked demo Space. The T3 checkpoint is loaded with multilingual vocabulary shape 2454 and S3 speech vocabulary shape 8194.
The demo combines these model-specific assets with the shared Chatterbox inference code and companion assets needed for end-to-end speech generation.