On the Evaluation of Neural Selective Prediction Methods for Natural Language Processing
Zhengyao Gu, Mark Hopkins
2023年份
1被引次数
3顶会引用
摘要
We provide a survey and empirical comparison of the state-of-the-art in neural selective classification for NLP tasks. We also provide a methodological blueprint, including a novel metric called refinement that provides a calibrated evaluation of confidence functions for selective prediction. Finally, we supply documented, open-source code to support the future development of selective prediction techniques.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- Teaching LLMs to Abstain across Languages via Multilingual FeedbackShangbin Feng, Weijia Shi, Yike Wang, Wenxuan Ding 等EMNLP 2024 · 被引用 4 次
- Fairness Beyond Performance: Revealing Reliability Disparities Across Groups in Legal NLPT. Y. S. S. Santosh, Irtiza ChowdhuryACL 2025
- From Input Perception to Predictive Insight: Modeling Model Blind Spots Before They Become ErrorsMaggie Mi, Aline Villavicencio, Nafise Sadat MoosaviEMNLP 2025
它引用的顶会 Paper4
- ViM: Out-Of-Distribution with Virtual-logit MatchingHaoqi Wang, Zhizhong Li, Litong Feng, Wayne ZhangCVPR 2022 · 被引用 227 次
- Selective Question Answering under Domain ShiftAmita Kamath, Robin Jia, Percy LiangACL 2020 · 被引用 121 次
- On the Inference Calibration of Neural Machine TranslationShuo Wang, Zhaopeng Tu, Shuming Shi, Yang LiuACL 2020 · 被引用 66 次
- The Art of Abstention: Selective Prediction and Error Regularization for Natural Language ProcessingJi Xin, Raphael Tang, Yaoliang Yu, Jimmy LinACL 2021
相关 Paper
- Leveraging Data to Say No: Memory Augmented Plug-and-Play Selective PredictionAditya Sarkar, Yi Li, Jiacheng Cheng, Shlok Kumar Mishra 等ICLR 2026
- Confidence Estimation for Error Detection in Text-to-SQL SystemsOleg Somov, Elena TutubalinaAAAI 2025 · 被引用 11 次
- Know When to Abstain: Optimal Selective Classification with Likelihood RatiosAlvin Heng, Harold SohICLR 2026 · 被引用 7 次
- A Model-Agnostic Heuristics for Selective ClassificationAndrea Pugnana, Salvatore RuggieriAAAI 2023 · 被引用 12 次
- Confidence Calibration of Classifiers with Many ClassesAdrien Le-Coz, Stéphane Herbin, Faouzi AdjedNeurIPS 2024 · 被引用 21 次
