Rethinking Calibration of Deep Neural Networks: Do Not Be Afraid of Overconfidence
Deng-Bao Wang, Lei Feng, Min-Ling Zhang
摘要
Capturing accurate uncertainty quantification of the predictions from deep neural networks is important in many real-world decision-making applications. A reliable predictor is expected to be accurate when it is confident about its predictions and indicate high uncertainty when it is likely to be inaccurate. However, modern neural networks have been found to be poorly calibrated, primarily in the direction of overconfidence. In recent years, there is a surge of research on model calibration by leveraging implicit or explicit regularization techniques during training, which achieve well calibration performance by avoiding overconfident outputs. In our study, we empirically found that despite the predictions obtained from these regularized models are better calibrated, they suffer from not being as calibratable, namely, it is harder to further calibrate these predictions with post-hoc calibration methods like temperature scaling and histogram binning. We conduct a series of empirical studies showing that overconfidence may not hurt final calibration performance if post-hoc calibration is allowed, rather, the penalty of confident outputs will compress the room of potential improvement in post-hoc calibration phase. Our experimental findings point out a new direction to improve calibration of DNNs by considering main training and post-hoc calibration as a unified framework.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper43
- Mitigating Neural Network Overconfidence with Logit NormalizationHongxin Wei, Renchunzi Xie, Hao Cheng, Lei Feng 等ICML 2022 · 被引用 386 次
- AdaFocal: Calibration-aware Adaptive Focal LossArindam Ghosh, Thomas Schaaf, Matthew GormleyNeurIPS 2022 · 被引用 71 次
- Dual Focal Loss for CalibrationLinwei Tao, Minjing Dong, Chang XuICML 2023 · 被引用 56 次
- Towards Improving Calibration in Object Detection Under Domain ShiftMuhammad Akhtar Munir, Muhammad Haris Khan, M. Saquib Sarfraz, Mohsen AliNeurIPS 2022 · 被引用 37 次
- Diffusion-Based Probabilistic Uncertainty Estimation for Active Domain AdaptationZhekai Du, Jingjing LiNeurIPS 2023 · 被引用 32 次
它引用的顶会 Paper7
- Calibrating Deep Neural Networks using Focal LossJishnu Mukhoti, Viveka Kulharia, Amartya Sanyal, Stuart Golodetz 等NeurIPS 2020 · 被引用 674 次
- Mix-n-Match : Ensemble and Compositional Methods for Uncertainty Calibration in Deep LearningJize Zhang, Bhavya Kailkhura, Thomas Yong-Jin HanICML 2020 · 被引用 276 次
- Improving model calibration with accuracy versus uncertainty optimizationRanganath Krishnan, Omesh TickooNeurIPS 2020 · 被引用 217 次
- Calibration of Neural Networks using SplinesKartik Gupta, Amir Rahimi, Thalaiyasingam Ajanthan, Thomas Mensink 等ICLR 2021 · 被引用 128 次
- Intra Order-preserving Functions for Calibration of Multi-Class Neural NetworksAmir Rahimi, Amirreza Shaban, Ching-An Cheng, Richard Hartley 等NeurIPS 2020 · 被引用 96 次
相关 Paper
- Post-Hoc Uncertainty Calibration for Domain Drift ScenariosChristian Tomani, Sebastian Gruber, Muhammed Ebrar Erdem, Daniel Cremers 等CVPR 2021
- Feature Clipping for Uncertainty CalibrationLinwei Tao, Minjing Dong, Chang XuAAAI 2025 · 被引用 6 次
- Uncertainty Quantification and Deep EnsemblesRahul Rahaman, Alexandre H. ThiéryNeurIPS 2021 · 被引用 250 次
- Parametric ρ-Norm Scaling CalibrationSiyuan Zhang, Linbo XieAAAI 2025
- Beyond In-Domain Scenarios: Robust Density-Aware CalibrationChristian Tomani, Futa Kai Waseda, Yuesong Shen, Daniel CremersICML 2023 · 被引用 16 次
