33 repos
Libraries, models, and tools for automatic speech recognition (ASR), speaker diarization, and audio processing tasks. The cluster spans multiple deployment contexts—from on-device inference using CoreML and ONNX Runtime to browser-based implementations with transformers.js—making these resources useful for building speech-enabled applications across platforms. Central repositories include optimized model variants like GPA and specialized audio models such as LFM2.5 and Raon-SpeechChat, reflecting a focus on making speech AI practical and accessible across different hardware and runtime constraints.