Automatic Speech Recognition & Audio Processing

18 repos

Speech-to-text and audio processing systems built primarily with Python and deep learning frameworks. The cluster centers on ASR model implementations and optimizations, including fine-tuned variants of Whisper and other speech recognition architectures. You'll find model compression techniques (distillation, quantization), language-specific adaptations, and inference optimizations across inference frameworks like JAX, alongside foundational audio processing and feature extraction tools.

Python · 5
Jupyter Notebook · 1
speech-recognition ·20,955
speech-to-text ·16,732
deep-learning ·16,502
audio ·15,937
asr ·12,065
speech-processing ·11,857
audio-processing ·11,821
language-model ·11,814
speaker-diarization ·11,814
huggingface ·11,814