11 repos
Libraries, models, and frameworks for automatic speech recognition (ASR) and audio processing, with a focus on PyTorch-based implementations. The cluster centers on end-to-end neural architectures like CTC, RNN-T, and TDT for converting speech to text, with several pre-trained Parakeet model variants at different scales. Developers here will find model implementations, training pipelines, and inference tools for building speech-to-text systems.