Energy-Based Open-World Uncertainty Modeling for Confidence Calibration
Yezhen Wang, Bo Li, Tong Che, Kaiyang Zhou, Ziwei Liu, Dongsheng Li
摘要
Confidence calibration is of great importance to the reliability of decisions made by machine learning systems. However, discriminative classifiers based on deep neural networks are often criticized for producing overconfident predictions that fail to reflect the true correctness likelihood of classification accuracy. We argue that such an inability to model uncertainty is mainly caused by the closed-world nature in softmax: a model trained by the cross-entropy loss will be forced to classify input into one of K pre-defined categories with high probability. To address this problem, we for the first time propose a novel K+1-way softmax formulation, which incorporates the modeling of open-world uncertainty as the extra dimension. To unify the learning of the original K-way classification task and the extra dimension that models uncertainty, we 1) propose a novel energy-based objective function, and moreover, 2) theoretically prove that optimizing such an objective essentially forces the extra dimension to capture the marginal data distribution. Extensive experiments show that our approach, Energy-based Open-World Softmax (EOW-Softmax), is superior to existing state-of-the-art methods in improving confidence calibration.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper27
- Negative Label Guided OOD Detection with Pretrained Vision-Language ModelsXue Jiang, Feng Liu, Zhen Fang, Hong Chen 等ICLR 2024 · 被引用 73 次
- Learning with Mixture of Prototypes for Out-of-Distribution DetectionHaodong Lu, Dong Gong, Shuo Wang, Jason Xue 等ICLR 2024 · 被引用 55 次
- Geometric Anchor Correspondence Mining with Uncertainty Modeling for Universal Domain AdaptationLiang Chen, Yihang Lou, Jianzhong He, Tao Bai 等CVPR 2022 · 被引用 46 次
- Watermarking for Out-of-distribution DetectionQizhou Wang, Feng Liu, Yonggang Zhang, Jing Zhang 等NeurIPS 2022 · 被引用 44 次
- Energy-based Latent Aligner for Incremental LearningK. J. Joseph, Salman Khan, Fahad Shahbaz Khan, Rao Muhammad Anwer 等CVPR 2022 · 被引用 35 次
它引用的顶会 Paper8
- Your classifier is secretly an energy based model and you should treat it like oneWill Grathwohl, Kuan-Chieh Wang, Jörn-Henrik Jacobsen, David Duvenaud 等ICLR 2020 · 被引用 643 次
- Generalized Energy Based ModelsMichael Arbel, Liang Zhou, Arthur GrettonICLR 2021 · 被引用 254 次
- On Calibration and Out-of-Domain GeneralizationYoav Wald, Amir Feder, Daniel Greenfeld, Uri ShalitNeurIPS 2021 · 被引用 184 次
- Learning Energy-Based Models by Diffusion Recovery LikelihoodRuiqi Gao, Yang Song, Ben Poole, Ying Nian Wu 等ICLR 2021 · 被引用 144 次
- Your GAN is Secretly an Energy-based Model and You Should Use Discriminator Driven Latent SamplingTong Che, Ruixiang Zhang, Jascha Sohl-Dickstein, Hugo Larochelle 等NeurIPS 2020 · 被引用 128 次
相关 Paper
- Energy-based Out-of-distribution DetectionWeitang Liu, Xiaoyun Wang, John D. Owens, Yixuan LiNeurIPS 2020 · 被引用 2,213 次
- Being Bayesian about Categorical ProbabilityTaejong Joo, Uijung Chung, Min-Gwan SeoICML 2020 · 被引用 69 次
- Confidence-Aware Learning for Deep Neural NetworksJooyoung Moon, Jihyo Kim, Younghak Shin, Sangheum HwangICML 2020 · 被引用 184 次
- Dual Energy-Based Model with Open-World Uncertainty Estimation for Out-of-distribution DetectionQi Chen, Hu DingCVPR 2025
- OpenMix: Exploring Outlier Samples for Misclassification DetectionFei Zhu, Zhen Cheng, Xu-Yao Zhang, Cheng-Lin LiuCVPR 2023
