Multilingual Speech Recognition Models

21 repos

Pre-trained multilingual automatic speech recognition models built on wav2vec 2.0 and transformer architectures, enabling speech-to-text across diverse languages including Hungarian, Finnish, Persian, Chinese, Arabic, and Greek. The cluster primarily consists of fine-tuned model repositories that apply the XLSR-53 (cross-lingual speech representations) framework to language-specific datasets, leveraging PyTorch and the Hugging Face transformers library for production speech recognition applications.

Python · 1
audio ·1,049
automatic-speech-recognition ·1,049
transformers ·1,049
speech ·1,045
endpoints_compatible ·586
pytorch ·586
wav2vec2 ·581
model-index ·530
jax ·529
xlsr-fine-tuning-week ·523