Explaining (Sarcastic) Utterances to Enhance Affect Understanding in Multimodal Dialogues
Shivani Kumar, Ishani Mondal, Md. Shad Akhtar, Tanmoy Chakraborty
摘要
Conversations emerge as the primary media for exchanging ideas and conceptions. From the listener's perspective, identifying various affective qualities, such as sarcasm, humour, and emotions, is paramount for comprehending the true connotation of the emitted utterance. However, one of the major hurdles faced in learning these affect dimensions is the presence of figurative language viz. irony, metaphor, or sarcasm. We hypothesize that any detection system constituting the exhaustive and explicit presentation of the emitted utterance would improve the overall comprehension of the dialogue. To this end, we explore the task of Sarcasm Explanation in Dialogues that aims to unfold the hidden irony behind sarcastic utterances. We propose MOSES, a deep neural network, which takes a multimodal (sarcastic) dialogue instance as an input and generates a natural language sentence as its explanation. Subsequently, we leverage the generated explanation for various natural language understanding tasks in a conversational dialogue setup, such as sarcasm detection, humour identification, and emotion recognition. Our evaluation shows that MOSES outperforms the state-of-the-art system for SED by an average of ∼ 2% on different evaluation metrics, such as ROUGE, BLEU, and METEOR. Further, we observe that leveraging the generated explanation advances three downstream tasks for affect classification -an average improvement of ∼ 14% F1-score in the sarcasm detection task and and ∼ 2% in the humour identification and emotion recognition task. We also perform extensive analyses to assess the quality of the results.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- Deciphering Cognitive Distortions in Patient-Doctor Mental Health Conversations: A Multimodal LLM-Based Detection and Reasoning FrameworkGopendra Vikram Singh, Sai Vemulapalli, Mauajama Firdaus, Asif EkbalEMNLP 2024 · 被引用 3 次
- VAGUE: Visual Contexts Clarify Ambiguous ExpressionsHeejeong Nam, Jinwoo Ahn, Keummin Ka, Jiwan Chung 等ICCV 2025 · 被引用 1 次
- MuVaC: A Variational Causal Framework for Multimodal Sarcasm Understanding in DialoguesDiandian Guo, Fangfang Yuan, Cong Cao, Xixun Lin 等WWW 2026
它引用的顶会 Paper5
- BART: Denoising Sequence-to-Sequence Pre-training for Natural Language Generation, Translation, and ComprehensionMike Lewis, Yinhan Liu, Naman Goyal, Marjan Ghazvininejad 等ACL 2020 · 被引用 1,224 次
- Sentiment and Emotion help Sarcasm? A Multi-task Learning Framework for Multi-Modal Sarcasm, Sentiment and Emotion AnalysisDushyant Singh Chauhan, Dhanush S. R, Asif Ekbal, Pushpak BhattacharyyaACL 2020 · 被引用 131 次
- Humor Knowledge Enriched Transformer for Understanding Multimodal HumorMd. Kamrul Hasan, Sangwu Lee, Wasifur Rahman, Amir Zadeh 等AAAI 2021 · 被引用 98 次
- R^3: Reverse, Retrieve, and Rank for Sarcasm Generation with Commonsense KnowledgeTuhin Chakrabarty, Debanjan Ghosh, Smaranda Muresan, Nanyun PengACL 2020 · 被引用 58 次
- When did you become so smart, oh wise one?! Sarcasm Explanation in Multi-modal Multi-party DialoguesShivani Kumar, Atharva Kulkarni, Md. Shad Akhtar, Tanmoy ChakrabortyACL 2022 · 被引用 54 次
相关 Paper
- Nice Perfume. How Long Did You Marinate in It? Multimodal Sarcasm ExplanationPoorav Desai, Tanmoy Chakraborty, Md. Shad AkhtarAAAI 2022 · 被引用 49 次
- Your tone speaks louder than your face! Modality Order Infused Multi-modal Sarcasm DetectionMohit Tomar, Abhisek Tiwari, Tulika Saha, Sriparna SahaACM MM 2023 · 被引用 15 次
- Multi-source Semantic Graph-based Multimodal Sarcasm Explanation GenerationLiqiang Jing, Xuemeng Song, Kun Ouyang, Mengzhao Jia 等ACL 2023 · 被引用 17 次
- Well, Now We Know! Unveiling Sarcasm: Initiating and Exploring Multimodal Conversations with ReasoningGopendra Vikram Singh, Mauajama Firdaus, Dushyant Singh Chauhan, Asif Ekbal 等AAAI 2024 · 被引用 6 次
- Incorporating Communication Style and Interaction of Speakers for Sarcasm Explanation in DialogueYuqing Li, Wenyuan Zhang, Zheng Lin, Guoxuan Ding 等SIGIR 2025 · 被引用 1 次
