Speech Emotion Recognition with Transformers

6 repos

Deep learning models for detecting emotional attributes (valence, arousal, dominance) and categorical emotions from audio speech using transformer-based architectures like WavLM and wav2vec2. These repositories implement baseline systems and fine-tuned models for the SER Odyssey challenge and related emotion recognition tasks, leveraging PyTorch and pre-trained speech representations to classify emotional content from voice signals.

audio-classification ·201
pytorch ·201
safetensors ·201
audio ·201
speech ·201
transformers ·201
emotion-recognition ·201
endpoints_compatible ·173
wav2vec2 ·173
wavlm ·28