The Art of Abstention: Selective Prediction and Error Regularization for Natural Language Processing
Ji Xin, Raphael Tang, Yaoliang Yu, Jimmy Lin
摘要
In selective prediction, a classifier is allowed to abstain from making predictions on lowconfidence examples. Though this setting is interesting and important, selective prediction has rarely been examined in natural language processing (NLP) tasks. To fill this void in the literature, we study in this paper selective prediction for NLP, comparing different models and confidence estimators. We further propose a simple error regularization trick that improves confidence estimation without substantially increasing the computation budget. We show that recent pre-trained transformer models simultaneously improve both model accuracy and confidence estimation effectiveness. We also find that our proposed regularization improves confidence estimation and can be applied to other relevant scenarios, such as using classifier cascades for accuracyefficiency trade-offs. Source code for this paper can be found at https://github.com/ castorini/transformers-selective .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper19
- Uncertainty Estimation of Transformer Predictions for Misclassification DetectionArtem Vazhentsev, Gleb Kuzmin, Artem Shelmanov, Akim Tsvigun 等ACL 2022 · 被引用 59 次
- Overcoming Common Flaws in the Evaluation of Selective Classification SystemsJeremias Traub, Till J. Bungert, Carsten T. Lüth, Michael Baumgartner 等NeurIPS 2024 · 被引用 44 次
- Learning to Reject with a Fixed Predictor: Application to DecontextualizationChristopher Mohri, Daniel Andor, Eunsol Choi, Michael Collins 等ICLR 2024 · 被引用 32 次
- Efficient Hallucination Detection for LLMs Using Uncertainty-Aware Attention HeadsArtem Vazhentsev, Lyudmila Rvanova, Gleb Kuzmin, Ekaterina Fadeeva 等ICML 2026 · 被引用 16 次
- Can NLI Provide Proper Indirect Supervision for Low-resource Biomedical Relation Extraction?Jiashu Xu, Mingyu Derek Ma, Muhao ChenACL 2023 · 被引用 14 次
它引用的顶会 Paper5
- ALBERT: A Lite BERT for Self-supervised Learning of Language RepresentationsZhenzhong Lan, Mingda Chen, Sebastian Goodman, Kevin Gimpel 等ICLR 2020 · 被引用 7,418 次
- BERT Loses Patience: Fast and Robust Inference with Early ExitWangchunshu Zhou, Canwen Xu, Tao Ge, Julian J. McAuley 等NeurIPS 2020 · 被引用 473 次
- Selective Question Answering under Domain ShiftAmita Kamath, Robin Jia, Percy LiangACL 2020 · 被引用 121 次
- On the Inference Calibration of Neural Machine TranslationShuo Wang, Zhaopeng Tu, Shuming Shi, Yang LiuACL 2020 · 被引用 66 次
- The Right Tool for the Job: Matching Model and Instance ComplexitiesRoy Schwartz, Gabriel Stanovsky, Swabha Swayamdipta, Jesse Dodge 等ACL 2020 · 被引用 4 次
相关 Paper
- On the Evaluation of Neural Selective Prediction Methods for Natural Language ProcessingZhengyao Gu, Mark HopkinsACL 2023 · 被引用 1 次
- Post-Abstention: Towards Reliably Re-Attempting the Abstained Instances in QANeeraj Varshney, Chitta BaralACL 2023 · 被引用 3 次
- Consistent Accelerated Inference via Confident Adaptive TransformersTal Schuster, Adam Fisch, Tommi S. Jaakkola, Regina BarzilayEMNLP 2021 · 被引用 30 次
- Confidence Estimation for Error Detection in Text-to-SQL SystemsOleg Somov, Elena TutubalinaAAAI 2025 · 被引用 11 次
- MASKER: Masked Keyword Regularization for Reliable Text ClassificationSeung Jun Moon, Sangwoo Mo, Kimin Lee, Jaeho Lee 等AAAI 2021 · 被引用 39 次
