25 repos
Lightweight, optimized speech recognition, text-to-speech, and voice processing models designed for on-device inference on mobile and edge hardware. This cluster focuses on compact model architectures (typically under 2B parameters) that enable real-time audio applications without cloud dependencies, featuring streaming ASR systems, voice synthesis, and multimodal voice models from vendors like NVIDIA, Alibaba, and Apple-compatible platforms.