37 repos
Neural text-to-speech models and implementations for converting written text into spoken audio. The cluster centers on deep learning TTS architectures, audio generation pipelines, and model checkpoints (including the Irodori 500M and DOTS SOAR variants), with emphasis on practical implementations using modern formats like safetensors for efficient model distribution. Repositories here cover both model training frameworks and inference tools for building voice synthesis applications.