ICASSP 2023accepted0 citations
Learning Robust Self-Attention Features for Speech Emotion Recognition with Label-Adaptive Mixup
Lei Kang, Lichao Zhang, Dazhi Jiang
Abstract
Speech Emotion Recognition (SER) is to recognize human emotions in a natural verbal interaction scenario with machines, which is considered as a challenging problem due to the ambiguous human emotions. Despite the recent progress in SER, state-of-the-art models struggle to achieve a satisfactory performance. We propose a self-attention based method with combined use of label-adaptive mixup and center loss. By adapting label probabilities in mixup and fitting center loss to the mixup training scheme, our proposed method achieves a superior performance to the state-of-the-art methods.
BibTeX
@inproceedings{icassp2023_learningrobustse,
title = {Learning Robust Self-Attention Features for Speech Emotion Recognition with Label-Adaptive Mixup},
author = {Lei Kang and Lichao Zhang and Dazhi Jiang},
booktitle = {ICASSP 2023},
year = {2023}
}