Towards Trustworthy Explanation: On Causal Rationalization
Wenbo Zhang, Tong Wu, Yunlong Wang, Yong Cai, Hengrui Cai
Abstract
With recent advances in natural language processing, rationalization becomes an essential self-explaining diagram to disentangle the black box by selecting a subset of input texts to account for the major variation in prediction. Yet, existing association-based approaches on rationalization cannot identify true rationales when two or more snippets are highly inter-correlated and thus provide a similar contribution to prediction accuracy, so-called spuriousness. To address this limitation, we novelly leverage two causal desiderata, non-spuriousness and efficiency, into rationalization from the causal inference perspective. We formally define a series of probabilities of causation based on a newly proposed structural causal model of rationalization, with its theoretical identification established as the main component of learning necessary and sufficient rationales. The superior performance of the proposed causal rationalization is demonstrated on real-world review and medical datasets with extensive experiments compared to state-of-the-art methods.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext cbb4c00a-f969-4a6d-9e3e-f71e66a52747Cited by top-tier papers10
- D-Separation for Causal Self-ExplanationWei Liu, Jun Wang, Haozhao Wang, Ruixuan Li et al.NeurIPS 2023 · 29 citations
- On Learning Necessary and Sufficient Causal GraphsHengrui Cai, Yixin Wang, Michael I. Jordan, Rui SongNeurIPS 2023 · 19 citations
- Is the MMI Criterion Necessary for Interpretability? Degenerating Non-causal Features to Plain Noise for Self-RationalizationWei Liu, Zhiying Deng, Zhongyu Niu, Jun Wang et al.NeurIPS 2024 · 17 citations
- Enhancing the Rationale-Input Alignment for Self-explaining RationalizationWei Liu, Haozhao Wang, Jun Wang, Zhiying Deng et al.ICDE 2024 · 6 citations
- GNN Explanations that do not Explain and How to find ThemSteve Azzolin, Stefano Teso, Bruno Lepri, Andrea Passerini et al.ICLR 2026 · 4 citations
Builds on16
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah et al.NeurIPS 2020 · 64,255 citations
- Invariant RationalizationShiyu Chang, Yang Zhang, Mo Yu, Tommi S. JaakkolaICML 2020 · 232 citations
- Explaining Black-Box Algorithms Using Probabilistic Contrastive CounterfactualsSainyam Galhotra, Romila Pradhan, Babak SalimiSIGMOD 2021 · 85 citations
- Understanding Interlocking Dynamics of Cooperative RationalizationMo Yu, Yang Zhang, Shiyu Chang, Tommi S. JaakkolaNeurIPS 2021 · 52 citations
- UNIREX: A Unified Learning Framework for Language Model Rationale ExtractionAaron Chan, Maziar Sanjabi, Lambert Mathias, Liang Tan et al.ICML 2022 · 48 citations
Related papers
- Accurate and Explainable Recommendation via Review RationalizationSicheng Pan, Dongsheng Li, Hansu Gu, Tun Lu et al.WWW 2022 · 22 citations
- Interventional RationalizationLinan Yue, Qi Liu, Li Wang, Yanqing An et al.EMNLP 2023 · 8 citations
- Measuring Association Between Labels and Free-Text RationalesSarah Wiegreffe, Ana Marasovic, Noah A. SmithEMNLP 2021 · 12 citations
- Learning Robust Rationales for Model Explainability: A Guidance-Based ApproachShuaibo Hu, Kui YuAAAI 2024 · 11 citations
- SPECTRA: Sparse Structured Text RationalizationNuno Miguel Guerreiro, André F. T. MartinsEMNLP 2021 · 1 citation
