Handling Ambiguity in Emotion: From Out-of-Domain Detection to Distribution Estimation
Wen Wu, Bo Li, Chao Zhang, Chung-Cheng Chiu, Qiujia Li, Junwen Bai, Tara N. Sainath, Philip C. Woodland
Abstract
The subjective perception of emotion leads to inconsistent labels from human annotators. Typically, utterances lacking majority-agreed labels are excluded when training an emotion classifier, which cause problems when encountering ambiguous emotional expressions during testing. This paper investigates three methods to handle ambiguous emotion. First, we show that incorporating utterances without majorityagreed labels as an additional class in the classifier reduces the classification performance of the other emotion classes. Then, we propose detecting utterances with ambiguous emotions as out-of-domain samples by quantifying the uncertainty in emotion classification using evidential deep learning. This approach retains the classification accuracy while effectively detects ambiguous emotion expressions. Furthermore, to obtain fine-grained distinctions among ambiguous emotions, we propose representing emotion as a distribution instead of a single class label. The task is thus re-framed from classification to distribution estimation where every individual annotation is taken into account, not just the majority opinion. The evidential uncertainty measure is extended to quantify the uncertainty in emotion distribution estimation. Experimental results on the IEMOCAP and CREMA-D datasets demonstrate the superior capability of the proposed method in terms of majority class prediction, emotion distribution estimation, and uncertainty estimation.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers1
Ask how each one uses itBuilds on3
- wav2vec 2.0: A Framework for Self-Supervised Learning of Speech RepresentationsAlexei Baevski, Yuhao Zhou, Abdelrahman Mohamed, Michael AuliNeurIPS 2020 · 9,451 citations
- data2vec: A General Framework for Self-supervised Learning in Speech, Vision and LanguageAlexei Baevski, Wei-Ning Hsu, Qiantong Xu, Arun Babu et al.ICML 2022 · 1,123 citations
- Large Language Models are Efficient Learners of Noise-Robust Speech RecognitionYuchen Hu, Chen Chen, Chao-Han Huck Yang, Ruizhe Li et al.ICLR 2024 · 41 citations
Related papers
- Estimating the Uncertainty in Emotion Attributes using Deep Evidential RegressionWen Wu, Chao Zhang, Philip C. WoodlandACL 2023 · 6 citations
- Semi-supervised Multi-modal Emotion Recognition with Cross-Modal Distribution MatchingJingjun Liang, Ruichen Li, Qin JinACM MM 2020 · 67 citations
- Dive Into Ambiguity: Latent Distribution Mining and Pairwise Uncertainty Estimation for Facial Expression RecognitionJiahui She, Yibo Hu, Hailin Shi, Jun Wang et al.CVPR 2021
- Is Epistemic Uncertainty Faithfully Represented by Evidential Deep Learning Methods?Mira Jürgens, Nis Meinert, Viktor Bengs, Eyke Hüllermeier et al.ICML 2024 · 35 citations
- Uncertainty-Aware Reliable Text ClassificationYibo Hu, Latifur KhanKDD 2021 · 25 citations
