Evaluating and Characterizing Human Rationales
Samuel Carton, Anirudh Rathore, Chenhao Tan
Abstract
Two main approaches for evaluating the quality of machine-generated rationales are: 1) using human rationales as a gold standard; and 2) automated metrics based on how rationales affect model behavior. An open question, however, is how human rationales fare with these automatic metrics. Analyzing a variety of datasets and models, we find that human rationales do not necessarily perform well on these metrics. To unpack this finding, we propose improved metrics to account for modeldependent baseline performance. We then propose two methods to further characterize rationale quality, one based on model retraining and one on using "fidelity curves" to reveal properties such as irrelevance and redundancy. Our work leads to actionable suggestions for evaluating and characterizing rationales.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 4977e2fb-e7b9-4e7c-aa52-11971506c230Cited by top-tier papers10
- UNIREX: A Unified Learning Framework for Language Model Rationale ExtractionAaron Chan, Maziar Sanjabi, Lambert Mathias, Liang Tan et al.ICML 2022 · 48 citations
- Towards Interactivity and Interpretability: A Rationale-based Legal Judgment Prediction FrameworkYiquan Wu, Yifei Liu, Weiming Lu, Yating Zhang et al.EMNLP 2022 · 33 citations
- Flexible Instance-Specific Rationalization of NLP ModelsGeorge Chrysostomou, Nikolaos AletrasAAAI 2022 · 17 citations
- Unifying Model Explainability and Robustness for Joint Text Classification and Rationale ExtractionDongfang Li, Baotian Hu, Qingcai Chen, Tujie Xu et al.AAAI 2022 · 16 citations
- Does Self-Rationalization Improve Robustness to Spurious Correlations?Alexis Ross, Matthew E. Peters, Ana MarasovicEMNLP 2022 · 4 citations
Builds on4
- Manipulating and Measuring Model InterpretabilityForough Poursabzi-Sangdeh, Daniel G. Goldstein, Jake M. Hofman, Jennifer Wortman Vaughan et al.CHI 2021 · 663 citations
- "Why is 'Chicago' deceptive?" Towards Building Model-Driven Tutorials for HumansVivian Lai, Han Liu, Chenhao TanCHI 2020 · 113 citations
- ERASER: A Benchmark to Evaluate Rationalized NLP ModelsJay DeYoung, Sarthak Jain, Nazneen Fatema Rajani, Eric P. Lehman et al.ACL 2020 · 36 citations
- An Information Bottleneck Approach for Controlling Conciseness in Rationale ExtractionBhargavi Paranjape, Mandar Joshi, John Thickstun, Hannaneh Hajishirzi et al.EMNLP 2020 · 13 citations
Related papers
- Are Machine Rationales (Not) Useful to Humans? Measuring and Improving Human Utility of Free-text RationalesBrihi Joshi, Ziyi Liu, Sahana Ramnath, Aaron Chan et al.ACL 2023 · 6 citations
- RORA: Robust Free-Text Rationale EvaluationZhengping Jiang, Yining Lu, Hanjie Chen, Daniel Khashabi et al.ACL 2024
- REV: Information-Theoretic Evaluation of Free-Text RationalesHanjie Chen, Faeze Brahman, Xiang Ren, Yangfeng Ji et al.ACL 2023 · 14 citations
- Measuring Association Between Labels and Free-Text RationalesSarah Wiegreffe, Ana Marasovic, Noah A. SmithEMNLP 2021 · 12 citations
- Making a (Counterfactual) Difference One Rationale at a TimeMitchell Plyler, Michael Green, Min ChiNeurIPS 2021 · 12 citations
