18 repos
Speech-to-text and audio processing systems built primarily with Python and deep learning frameworks. The cluster centers on ASR model implementations and optimizations, including fine-tuned variants of Whisper and other speech recognition architectures. You'll find model compression techniques (distillation, quantization), language-specific adaptations, and inference optimizations across inference frameworks like JAX, alongside foundational audio processing and feature extraction tools.