On-Device Speech & Audio AI Models

25 repos

Lightweight, optimized speech recognition, text-to-speech, and voice processing models designed for on-device inference on mobile and edge hardware. This cluster focuses on compact model architectures (typically under 2B parameters) that enable real-time audio applications without cloud dependencies, featuring streaming ASR systems, voice synthesis, and multimodal voice models from vendors like NVIDIA, Alibaba, and Apple-compatible platforms.

Swift · 1
core-ai ·1,898
apple-silicon ·1,888
coreml ·1,885
super-resolution ·1,874
object-detection ·1,874
mlpackage ·1,871
coremltools ·1,871
deep-learning ·1,870
ios ·1,870
gan ·1,870