Self-training with Few-shot Rationalization
Meghana Moorthy Bhat, Alessandro Sordoni, Subhabrata Mukherjee
Abstract
While pre-trained language models have obtained state-of-the-art performance for several natural language understanding tasks, they are quite opaque in terms of their decision-making process. While some recent works focus on rationalizing neural predictions by highlighting salient concepts in text as justifications or rationales, they rely on thousands of labeled training examples for both task labels as well as annotated rationales for every instance. Such extensive large-scale annotations are infeasible to obtain for many tasks. To this end, we develop a multi-task teacher-student framework based on self-training language models with limited task-specific labels and rationales, and judicious sample selection to learn from informative pseudo-labeled examples 1 . We study several characteristics of what constitutes a good rationale and demonstrate that the neural model performance can be significantly improved by making it aware of its rationalized predictions particularly in low-resource settings. Extensive experiments in several benchmark datasets demonstrate the effectiveness of our approach.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext d38bade9-416e-4073-be07-e0f83fbcee3aCited by top-tier papers2
- Does Self-Rationalization Improve Robustness to Spurious Correlations?Alexis Ross, Matthew E. Peters, Ana MarasovicEMNLP 2022 · 4 citations
- DuNST: Dual Noisy Self Training for Semi-Supervised Controllable Text GenerationYuxi Feng, Xiaoyuan Yi, Xiting Wang, Laks V. S. Lakshmanan et al.ACL 2023 · 2 citations
Builds on5
- Uncertainty-aware Self-training for Few-shot Text ClassificationSubhabrata Mukherjee, Ahmed Hassan AwadallahNeurIPS 2020 · 182 citations
- ERASER: A Benchmark to Evaluate Rationalized NLP ModelsJay DeYoung, Sarthak Jain, Nazneen Fatema Rajani, Eric P. Lehman et al.ACL 2020 · 36 citations
- An Information Bottleneck Approach for Controlling Conciseness in Rationale ExtractionBhargavi Paranjape, Mandar Joshi, John Thickstun, Hannaneh Hajishirzi et al.EMNLP 2020 · 13 citations
- Learning to Faithfully Rationalize by ConstructionSarthak Jain, Sarah Wiegreffe, Yuval Pinter, Byron C. WallaceACL 2020
- Self-Training With Noisy Student Improves ImageNet ClassificationQizhe Xie, Minh-Thang Luong, Eduard H. Hovy, Quoc V. LeCVPR 2020
Related papers
- PINTO: Faithful Language Reasoning Using Prompt-Generated RationalesPeifeng Wang, Aaron Chan, Filip Ilievski, Muhao Chen et al.ICLR 2023 · 21 citations
- Boosting Explainability through Selective Rationalization in Pre-trained Language ModelsLibing Yuan, Shuaibo Hu, Kui Yu, Le WuKDD 2025 · 1 citation
- Learning to Rationalize for Nonmonotonic Reasoning with Distant SupervisionFaeze Brahman, Vered Shwartz, Rachel Rudinger, Yejin ChoiAAAI 2021 · 46 citations
- Learn from A Rationalist: Distilling Intermediate Interpretable RationalesJiayi Dai, Randy GoebelICML 2026
- Knowledge-Grounded Self-Rationalization via Extractive and Natural Language ExplanationsBodhisattwa Prasad Majumder, Oana Camburu, Thomas Lukasiewicz, Julian J. McAuleyICML 2022 · 40 citations
