32 repos
Python-based libraries, models, and frameworks for converting spoken audio into text using deep learning approaches. The cluster covers end-to-end ASR systems, model implementations (particularly with PyTorch and ONNX), training pipelines, and deployment-ready solutions. Includes both general-purpose transcription tools and specialized variants like multilingual and robust ASR models designed for real-world audio conditions.