29 repos
Tools and models for synthesizing speech from text and converting between audio modalities, with a focus on efficient implementations across edge and mobile platforms. The cluster includes neural TTS engines, GGUF-quantized model variants for resource-constrained inference, and cross-platform implementations—particularly in Swift for iOS/macOS deployment. Central repos like VoxCPM, Soprano, and Kokoro represent compact language models optimized for real-time speech synthesis.