Detecting Misclassification Errors in Neural Networks with a Gaussian Process Model
Xin Qiu, Risto Miikkulainen
Abstract
As neural network classifiers are deployed in real-world applications, it is crucial that their failures can be detected reliably. One practical solution is to assign confidence scores to each prediction, then use these scores to filter out possible misclassifications. However, existing confidence metrics are not yet sufficiently reliable for this role. This paper presents a new framework that produces a quantitative metric for detecting misclassification errors. This framework, RED, builds an error detector on top of the base classifier and estimates uncertainty of the detection scores using Gaussian Processes. Experimental comparisons with other error detection methods on 125 UCI datasets demonstrate that this approach is effective. Further implementations on two probabilistic base classifiers and two large deep learning architecture in vision tasks further confirm that the method is robust and scalable. Third, an empirical analysis of RED with out-of-distribution and adversarial samples shows that the method can be used not only to detect errors but also to understand where they come from. RED can thereby be used to improve trustworthiness of neural network classifiers more broadly in the future.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 9b892232-7f70-48e6-983c-c1cff17bb8ceCited by top-tier papers4
- Adaptive Confidence Regularization for Multimodal Failure DetectionMoru Liu, Hao Dong, Olga Fink, Mario TrappCVPR 2026 · 1 citation
- Test Case Prioritization for DNNs via Neural Collapse InstabilityChunyu Liu, Mingyuan Li, Yang Li, Wenmin Li et al.ISSTA 2026
- Unveiling AI's Blind Spots: An Oracle for In-Domain, Out-of-Domain, and Adversarial ErrorsShuangpeng Han, Mengmi ZhangICML 2025
- REACH: Explicit Recovery Behavior for Diffusion PoliciesZundong Ke, Junlin Chen, Jiayi Zhu, Kuanhao Xia et al.CVPR 2026
Builds on3
- Simple and Principled Uncertainty Estimation with Deterministic Deep Learning via Distance AwarenessJeremiah Z. Liu, Zi Lin, Shreyas Padhy, Dustin Tran et al.NeurIPS 2020 · 604 citations
- Quantifying Point-Prediction Uncertainty in Neural Networks via Residual Estimation with an I/O KernelXin Qiu, Elliot Meyerson, Risto MiikkulainenICLR 2020 · 60 citations
- Distance-Based Learning from Errors for Confidence CalibrationChen Xing, Sercan Ömer Arik, Zizhao Zhang, Tomas PfisterICLR 2020 · 42 citations
Related papers
- Better Uncertainty Calibration via Proper Scores for Classification and BeyondSebastian G. Gruber, Florian BuettnerNeurIPS 2022 · 88 citations
- Adaptive Uncertainty Estimation via High-Dimensional Testing on Latent RepresentationsTsai Hor Chan, Kin Wai Lau, Jiajun Shen, Guosheng Yin et al.NeurIPS 2023 · 3 citations
- A Data-Driven Measure of Relative Uncertainty for Misclassification DetectionEduardo Dadalto Câmara Gomes, Marco Romanelli, Georg Pichler, Pablo PiantanidaICLR 2024 · 11 citations
- OpenMix: Exploring Outlier Samples for Misclassification DetectionFei Zhu, Zhen Cheng, Xu-Yao Zhang, Cheng-Lin LiuCVPR 2023
- Beyond Unimodal: Generalising Neural Processes for Multimodal Uncertainty EstimationMyong Chol Jung, He Zhao, Joanna Dipnall, Lan DuNeurIPS 2023 · 18 citations
