Speech Recognition & Whisper Inference

129 repos across 6 sub-areas

Libraries, tools, and applications for automatic speech-to-text conversion, with heavy focus on OpenAI's Whisper model and its deployment across multiple platforms. The cluster spans Python implementations for training and inference, Swift/iOS integrations, Rust-based optimizations for edge deployment, and specialized tools like speaker diarization and real-time transcription. Developers here will find production-ready inference engines, model optimization techniques, and cross-platform wrappers enabling Whisper deployment from cloud services to mobile and embedded devices.