Uncertainty Estimation of Transformer Predictions for Misclassification Detection
Artem Vazhentsev, Gleb Kuzmin, Artem Shelmanov, Akim Tsvigun, Evgenii Tsymbalov, Kirill Fedyanin, Maxim Panov, Alexander Panchenko, Gleb Gusev, Mikhail Burtsev, Manvel Avetisian, Leonid Zhukov
摘要
Uncertainty estimation (UE) of model predictions is a crucial step for a variety of tasks such as active learning, misclassification detection, adversarial attack detection, out-ofdistribution detection, etc. Most of the works on modeling the uncertainty of deep neural networks evaluate these methods on image classification tasks. Little attention has been paid to UE in natural language processing. To fill this gap, we perform a vast empirical investigation of state-of-the-art UE methods for Transformer models on misclassification detection in named entity recognition and text classification tasks and propose two computationally efficient modifications, one of which approaches or even outperforms computationally intensive methods 1 .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper10
- Shifting Attention to Relevance: Towards the Predictive Uncertainty Quantification of Free-Form Large Language ModelsJinhao Duan, Hao Cheng, Shiqi Wang, Alex Zavalny 等ACL 2024 · 被引用 28 次
- Hybrid Uncertainty Quantification for Selective Text Classification in Ambiguous TasksArtem Vazhentsev, Gleb Kuzmin, Akim Tsvigun, Alexander Panchenko 等ACL 2023 · 被引用 11 次
- Uncertainty Guided Label Denoising for Document-level Distant Relation ExtractionQi Sun, Kun Huang, Xiaocui Yang, Pengfei Hong 等ACL 2023 · 被引用 11 次
- Using Artificial Populations to Study Psychological Phenomena in Neural ModelsJesse Roberts, Kyle Moore, Drew Wilenzick, Douglas H. FisherAAAI 2024 · 被引用 8 次
- MARS: Meaning-Aware Response Scoring for Uncertainty Estimation in Generative LLMsYavuz Faruk Bakman, Duygu Nur Yaldiz, Baturalp Buyukates, Chenyang Tao 等ACL 2024 · 被引用 6 次
它引用的顶会 Paper10
- Deberta: decoding-Enhanced Bert with Disentangled AttentionPengcheng He, Xiaodong Liu, Jianfeng Gao, Weizhu ChenICLR 2021 · 被引用 3,729 次
- Simple and Principled Uncertainty Estimation with Deterministic Deep Learning via Distance AwarenessJeremiah Z. Liu, Zi Lin, Shreyas Padhy, Dustin Tran 等NeurIPS 2020 · 被引用 604 次
- ELECTRA: Pre-training Text Encoders as Discriminators Rather Than GeneratorsKevin Clark, Minh-Thang Luong, Quoc V. Le, Christopher D. ManningICLR 2020 · 被引用 541 次
- Pitfalls of In-Domain Uncertainty Estimation and Ensembling in Deep LearningArsenii Ashukha, Alexander Lyzhov, Dmitry Molchanov, Dmitry P. VetrovICLR 2020 · 被引用 354 次
- Revisiting Mahalanobis Distance for Transformer-Based Out-of-Domain DetectionAlexander Podolskiy, Dmitry Lipin, Andrey Bout, Ekaterina Artemova 等AAAI 2021 · 被引用 100 次
相关 Paper
- Are Data Augmentation Methods in Named Entity Recognition Applicable for Uncertainty Estimation?Wataru Hashimoto, Hidetaka Kamigaito, Taro WatanabeEMNLP 2024 · 被引用 1 次
- Transformer Uncertainty Estimation with Hierarchical Stochastic AttentionJiahuan Pei, Cheng Wang, György SzarvasAAAI 2022 · 被引用 33 次
- Uncertainty-Aware Reliable Text ClassificationYibo Hu, Latifur KhanKDD 2021 · 被引用 25 次
- The Art of Abstention: Selective Prediction and Error Regularization for Natural Language ProcessingJi Xin, Raphael Tang, Yaoliang Yu, Jimmy LinACL 2021
- Nonparametric Uncertainty Quantification for Single Deterministic Neural NetworkNikita Kotelevskii, Aleksandr Artemenkov, Kirill Fedyanin, Fedor Noskov 等NeurIPS 2022 · 被引用 52 次
