Towards Faithful Explanations: Boosting Rationalization with Shortcuts Discovery
Linan Yue, Qi Liu, Yichao Du, Li Wang, Weibo Gao, Yanqing An
Abstract
The remarkable success in neural networks provokes the selective rationalization. It explains the prediction results by identifying a small subset of the inputs sufficient to support them. Since existing methods still suffer from adopting the shortcuts in data to compose rationales and limited large-scale annotated rationales by human, in this paper, we propose a Shortcuts-fused Selective Rationalization (SSR) method, which boosts the rationalization by discovering and exploiting potential shortcuts. Specifically, SSR first designs a shortcuts discovery approach to detect several potential shortcuts. Then, by introducing the identified shortcuts, we propose two strategies to mitigate the problem of utilizing shortcuts to compose rationales. Finally, we develop two data augmentations methods to close the gap in the number of annotated rationales. Extensive experimental results on real-world datasets clearly validate the effectiveness of our proposed method. Code is released at https://github.com/yuelinan/codes-of-SSR .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers6
- Why and How LLMs Hallucinate: Connecting the Dots with Subsequence AssociationsYiyou Sun, Yu Gai, Lijie Chen, Abhilasha Ravichander et al.NeurIPS 2025 · 20 citations
- Is the MMI Criterion Necessary for Interpretability? Degenerating Non-causal Features to Plain Noise for Self-RationalizationWei Liu, Zhiying Deng, Zhongyu Niu, Jun Wang et al.NeurIPS 2024 · 17 citations
- GNN Explanations that do not Explain and How to find ThemSteve Azzolin, Stefano Teso, Bruno Lepri, Andrea Passerini et al.ICLR 2026 · 4 citations
- Do LLMs Overcome Shortcut Learning? An Evaluation of Shortcut Challenges in Large Language ModelsYu Yuan, Lili Zhao, Kai Zhang, Guangting Zheng et al.EMNLP 2024 · 4 citations
- Boosting Explainability through Selective Rationalization in Pre-trained Language ModelsLibing Yuan, Shuaibo Hu, Kui Yu, Le WuKDD 2025 · 1 citation
Builds on16
- Training language models to follow instructions with human feedbackLong Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida et al.NeurIPS 2022 · 24,707 citations
- Invariant RationalizationShiyu Chang, Yang Zhang, Mo Yu, Tommi S. JaakkolaICML 2020 · 232 citations
- Learning Invariant Graph Representations for Out-of-Distribution GeneralizationHaoyang Li, Ziwei Zhang, Xin Wang, Wenwu ZhuNeurIPS 2022 · 170 citations
- Understanding Interlocking Dynamics of Cooperative RationalizationMo Yu, Yang Zhang, Shiyu Chang, Tommi S. JaakkolaNeurIPS 2021 · 52 citations
- UNIREX: A Unified Learning Framework for Language Model Rationale ExtractionAaron Chan, Maziar Sanjabi, Lambert Mathias, Liang Tan et al.ICML 2022 · 48 citations
Related papers
- Interventional RationalizationLinan Yue, Qi Liu, Li Wang, Yanqing An et al.EMNLP 2023 · 8 citations
- Federated Self-Explaining GNNs with Anti-shortcut AugmentationsLinan Yue, Qi Liu, Weibo Gao, Ye Liu et al.ICML 2024 · 2 citations
- Learning Robust Rationales for Model Explainability: A Guidance-Based ApproachShuaibo Hu, Kui YuAAAI 2024 · 11 citations
- Distribution Matching for RationalizationYongfeng Huang, Yujun Chen, Yulun Du, Zhilin YangAAAI 2021 · 21 citations
- Difference-based Sample Selection for Federated Graph RationalizationLinan Yue, Weibo GaoWWW 2026
