Human Rationales as Attribution Priors for Explainable Stance Detection
Sahil Jayaram, Emily Allaway
摘要
As NLP systems become better at detecting opinions and beliefs from text, it is important to ensure not only that models are accurate but also that they arrive at their predictions in ways that align with human reasoning. In this work, we present a method for imparting human-like rationalization to a stance detection model using crowdsourced annotations on a small fraction of the training data. We show that in a data-scarce setting, our approach can improve the reasoning of a state-of-the-art classifierparticularly for inputs containing challenging phenomena such as sarcasm-at no cost in predictive performance. Furthermore, we demonstrate that attention weights surpass a leading attribution method in providing faithful explanations of our model's predictions, thus serving as a computationally cheap and reliable source of attributions for our model.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- Neglected Free Lunch - Learning Image Classifiers Using Annotation ByproductsDongyoon Han, Junsuk Choe, Seonghyeok Chun, John Joon Young Chung 等ICCV 2023 · 被引用 4 次
- Why Should This Article Be Deleted? Transparent Stance Detection in Multilingual Wikipedia Editor DiscussionsLucie-Aimée Kaffee, Arnav Arora, Isabelle AugensteinEMNLP 2023 · 被引用 3 次
- Leveraging Machine-Generated Rationales to Facilitate Social Meaning Detection in ConversationsRitam Dutt, Zhen Wu, Jiaxin Shi, Divyanshu Sheth 等ACL 2024 · 被引用 2 次
- SOCIAL SCAFFOLDS: A Generalization Framework for Social Understanding TasksRitam Dutt, Carolyn P. Rosé, Maarten SapEMNLP 2025
它引用的顶会 Paper5
- A Diagnostic Study of Explainability Techniques for Text ClassificationPepa Atanasova, Jakob Grue Simonsen, Christina Lioma, Isabelle AugensteinEMNLP 2020 · 被引用 158 次
- Enhancing Cross-target Stance Detection with Transferable Semantic-Emotion KnowledgeBowen Zhang, Min Yang, Xutao Li, Yunming Ye 等ACL 2020 · 被引用 115 次
- ERASER: A Benchmark to Evaluate Rationalized NLP ModelsJay DeYoung, Sarthak Jain, Nazneen Fatema Rajani, Eric P. Lehman 等ACL 2020 · 被引用 36 次
- Zero-Shot Stance Detection: A Dataset and Model using Generalized Topic RepresentationsEmily Allaway, Kathleen R. McKeownEMNLP 2020 · 被引用 7 次
- Cross-Domain Label-Adaptive Stance DetectionMomchil Hardalov, Arnav Arora, Preslav Nakov, Isabelle AugensteinEMNLP 2021 · 被引用 3 次
相关 Paper
- Human Attention Maps for Text Classification: Do Humans and Neural Networks Focus on the Same Words?Cansu Sen, Thomas Hartvigsen, Biao Yin, Xiangnan Kong 等ACL 2020 · 被引用 56 次
- S³-MSD: Large Vision-Language Model for Explainable and Generalizable Multi-modal Sarcasm DetectionZhihong Zhu, Fan Zhang, Yunyan Zhang, Jinghan Sun 等AAAI 2026
- Less is More: Attention Supervision with Counterfactuals for Text ClassificationSeungtaek Choi, Haeju Park, Jinyoung Yeo, Seung-won HwangEMNLP 2020 · 被引用 16 次
- Exploring the Efficacy of Automatically Generated Counterfactuals for Sentiment AnalysisLinyi Yang, Jiazheng Li, Padraig Cunningham, Yue Zhang 等ACL 2021
- Towards Multi-Modal Sarcasm Detection via Hierarchical Congruity Modeling with Knowledge EnhancementHui Liu, Wenya Wang, Haoliang LiEMNLP 2022 · 被引用 91 次
