6 repos
Deep learning models for detecting emotional attributes (valence, arousal, dominance) and categorical emotions from audio speech using transformer-based architectures like WavLM and wav2vec2. These repositories implement baseline systems and fine-tuned models for the SER Odyssey challenge and related emotion recognition tasks, leveraging PyTorch and pre-trained speech representations to classify emotional content from voice signals.