21 repos
Pre-trained multilingual automatic speech recognition models built on wav2vec 2.0 and transformer architectures, enabling speech-to-text across diverse languages including Hungarian, Finnish, Persian, Chinese, Arabic, and Greek. The cluster primarily consists of fine-tuned model repositories that apply the XLSR-53 (cross-lingual speech representations) framework to language-specific datasets, leveraging PyTorch and the Hugging Face transformers library for production speech recognition applications.