DARE: Disentanglement-Augmented Rationale Extraction
Linan Yue, Qi Liu, Yichao Du, Yanqing An, Li Wang, Enhong Chen
Abstract
Rationale extraction can be considered as a straightforward method of improving the model explainability, where rationales are a subsequence of the original inputs, and can be extracted to support the prediction results. Existing methods are mainly cascaded with the selector which extracts the rationale tokens, and the predictor which makes the prediction based on selected tokens. Since previous works fail to fully exploit the original input, where the information of non-selected tokens is ignored, in this paper, we propose a Disentanglement-Augmented Rationale Extraction (DARE) method, which encapsulates more information from the input to extract rationales. Specifically, it first disentangles the input into the rationale representations and the non-rationale ones, and then learns more comprehensive rationale representations for extracting by minimizing the mutual information (MI) between the two disentangled representations. Besides, to improve the performance of MI minimization, we develop a new MI estimator by exploring existing MI estimation methods. Extensive experimental results on three real-world datasets and simulation studies clearly validate the effectiveness of our proposed method. Code is released at https://github.com/yuelinan/DARE . ⇤ Corresponding Author 36th Conference on Neural Information Processing Systems (NeurIPS 2022). whole text as the input and generate the accurate but uninterpretable representations as most DNNs do to predict the result. Then, the guider utilizes the above representations to guide the predictor to yield more comprehensive task-related representations with an adversarial-based method. Since this method fails to utilize the information of the original text, where the non-rationale tokens are ignored, we argue that this "guidance pattern" can be further explored to improve the rationale extraction. After hearing, our court identified that the defendant and the victim had a dispute caused by trivial. The defendant slashed the victim with a knife, which caused a serious injury to the victim. Soon after, the defendant was under arrest by policeman...... selector predictor guider Charge: Crime of intentional injury After hearing, our court identified that the defendant and the victim had a dispute caused by trivial. The defendant slashed the victim with a knife, which caused a serious injury to the victim. Soon after, the defendant was under arrest by policeman...... After hearing, our court identified that the defendant and the victim had a dispute caused by trivial. The defendant slashed the victim with a knife, which caused a serious injury to the victim. Soon after, the defendant was under arrest by policeman......
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext b5fc1b98-23b2-4e1f-8f48-374e5973661fCited by top-tier papers15
- Leveraging Transferable Knowledge Concept Graph Embedding for Cold-Start Cognitive DiagnosisWeibo Gao, Hao Wang, Qi Liu, Fei Wang et al.SIGIR 2023 · 52 citations
- Cooperative Classification and Rationalization for Graph GeneralizationLinan Yue, Qi Liu, Ye Liu, Weibo Gao et al.WWW 2024 · 15 citations
- Learning Invariant Inter-pixel Correlations for Superpixel GenerationSen Xu, Shikui Wei, Tao Ruan, Lixin LiaoAAAI 2024 · 14 citations
- Collaborative Cognitive Diagnosis with Disentangled Representation Learning for Learner ModelingWeibo Gao, Qi Liu, Linan Yue, Fangzhou Yao et al.NeurIPS 2024 · 12 citations
- Learning Robust Rationales for Model Explainability: A Guidance-Based ApproachShuaibo Hu, Kui YuAAAI 2024 · 11 citations
Builds on20
- Big Bird: Transformers for Longer SequencesManzil Zaheer, Guru Guruganesh, Kumar Avinava Dubey, Joshua Ainslie et al.NeurIPS 2020 · 3,159 citations
- A Unified MRC Framework for Named Entity RecognitionXiaoya Li, Jingrong Feng, Yuxian Meng, Qinghong Han et al.ACL 2020 · 617 citations
- CLUB: A Contrastive Log-ratio Upper Bound of Mutual InformationPengyu Cheng, Weituo Hao, Shuyang Dai, Jiachang Liu et al.ICML 2020 · 512 citations
- Invariant RationalizationShiyu Chang, Yang Zhang, Mo Yu, Tommi S. JaakkolaICML 2020 · 232 citations
- Distinguish Confusing Law Articles for Legal Judgment PredictionNuo Xu, Pinghui Wang, Long Chen, Li Pan et al.ACL 2020 · 150 citations
Related papers
- Making a (Counterfactual) Difference One Rationale at a TimeMitchell Plyler, Michael Green, Min ChiNeurIPS 2021 · 12 citations
- Interventional RationalizationLinan Yue, Qi Liu, Li Wang, Yanqing An et al.EMNLP 2023 · 8 citations
- Learning from the Best: Rationalizing Predictions by Adversarial Information CalibrationLei Sha, Oana-Maria Camburu, Thomas LukasiewiczAAAI 2021 · 40 citations
- QUASER: Question Answering with Scalable Extractive RationalizationAsish Ghoshal, Srinivasan Iyer, Bhargavi Paranjape, Kushal Lakhotia et al.SIGIR 2022 · 2 citations
- D-Separation for Causal Self-ExplanationWei Liu, Jun Wang, Haozhao Wang, Ruixuan Li et al.NeurIPS 2023 · 29 citations
