The Art of Abstention: Selective Prediction and Error Regularization for Natural Language Processing
Ji Xin, Raphael Tang, Yaoliang Yu, Jimmy Lin
Abstract
In selective prediction, a classifier is allowed to abstain from making predictions on lowconfidence examples. Though this setting is interesting and important, selective prediction has rarely been examined in natural language processing (NLP) tasks. To fill this void in the literature, we study in this paper selective prediction for NLP, comparing different models and confidence estimators. We further propose a simple error regularization trick that improves confidence estimation without substantially increasing the computation budget. We show that recent pre-trained transformer models simultaneously improve both model accuracy and confidence estimation effectiveness. We also find that our proposed regularization improves confidence estimation and can be applied to other relevant scenarios, such as using classifier cascades for accuracyefficiency trade-offs. Source code for this paper can be found at https://github.com/ castorini/transformers-selective .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 83e3812c-4b20-4569-b461-b138576c65ecCited by top-tier papers19
- Uncertainty Estimation of Transformer Predictions for Misclassification DetectionArtem Vazhentsev, Gleb Kuzmin, Artem Shelmanov, Akim Tsvigun et al.ACL 2022 · 59 citations
- Overcoming Common Flaws in the Evaluation of Selective Classification SystemsJeremias Traub, Till J. Bungert, Carsten T. Lüth, Michael Baumgartner et al.NeurIPS 2024 · 44 citations
- Learning to Reject with a Fixed Predictor: Application to DecontextualizationChristopher Mohri, Daniel Andor, Eunsol Choi, Michael Collins et al.ICLR 2024 · 32 citations
- Efficient Hallucination Detection for LLMs Using Uncertainty-Aware Attention HeadsArtem Vazhentsev, Lyudmila Rvanova, Gleb Kuzmin, Ekaterina Fadeeva et al.ICML 2026 · 16 citations
- Can NLI Provide Proper Indirect Supervision for Low-resource Biomedical Relation Extraction?Jiashu Xu, Mingyu Derek Ma, Muhao ChenACL 2023 · 14 citations
Builds on5
- ALBERT: A Lite BERT for Self-supervised Learning of Language RepresentationsZhenzhong Lan, Mingda Chen, Sebastian Goodman, Kevin Gimpel et al.ICLR 2020 · 7,418 citations
- BERT Loses Patience: Fast and Robust Inference with Early ExitWangchunshu Zhou, Canwen Xu, Tao Ge, Julian J. McAuley et al.NeurIPS 2020 · 473 citations
- Selective Question Answering under Domain ShiftAmita Kamath, Robin Jia, Percy LiangACL 2020 · 121 citations
- On the Inference Calibration of Neural Machine TranslationShuo Wang, Zhaopeng Tu, Shuming Shi, Yang LiuACL 2020 · 66 citations
- The Right Tool for the Job: Matching Model and Instance ComplexitiesRoy Schwartz, Gabriel Stanovsky, Swabha Swayamdipta, Jesse Dodge et al.ACL 2020 · 4 citations
Related papers
- On the Evaluation of Neural Selective Prediction Methods for Natural Language ProcessingZhengyao Gu, Mark HopkinsACL 2023 · 1 citation
- Post-Abstention: Towards Reliably Re-Attempting the Abstained Instances in QANeeraj Varshney, Chitta BaralACL 2023 · 3 citations
- Consistent Accelerated Inference via Confident Adaptive TransformersTal Schuster, Adam Fisch, Tommi S. Jaakkola, Regina BarzilayEMNLP 2021 · 30 citations
- Confidence Estimation for Error Detection in Text-to-SQL SystemsOleg Somov, Elena TutubalinaAAAI 2025 · 11 citations
- MASKER: Masked Keyword Regularization for Reliable Text ClassificationSeung Jun Moon, Sangwoo Mo, Kimin Lee, Jaeho Lee et al.AAAI 2021 · 39 citations
