A Consistent and Differentiable Lp Canonical Calibration Error Estimator
Teodora Popordanoska, Raphael Sayer, Matthew B. Blaschko
Abstract
Calibrated probabilistic classifiers are models whose predicted probabilities can directly be interpreted as uncertainty estimates. It has been shown recently that deep neural networks are poorly calibrated and tend to output overconfident predictions. As a remedy, we propose a low-bias, trainable calibration error estimator based on Dirichlet kernel density estimates, which asymptotically converges to the true L p calibration error. This novel estimator enables us to tackle the strongest notion of multiclass calibration, called canonical (or distribution) calibration, while other common calibration methods are tractable only for top-label and marginal calibration. The computational complexity of our estimator is O(n 2 ), the convergence rate is O(n -1/2 ), and it is unbiased up to O(n -2 ), achieved by a geometric series debiasing scheme. In practice, this means that the estimator can be applied to small subsets of data, enabling efficient estimation and mini-batch updates. The proposed method has a natural choice of kernel, and can be used to generate consistent estimates of other quantities based on conditional expectation, such as the sharpness of a probabilistic classifier. Empirical results validate the correctness of our estimator, and demonstrate its utility in canonical calibration error estimation and calibration error regularized risk minimization. * Equal contribution † Most of this work was done while at KU Leuven 36th Conference on Neural Information Processing Systems (NeurIPS 2022).
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 11b22590-7f06-434d-9fdc-1a0f2c7fcfddCited by top-tier papers20
- Better Uncertainty Calibration via Proper Scores for Classification and BeyondSebastian G. Gruber, Florian BuettnerNeurIPS 2022 · 88 citations
- Dual Focal Loss for CalibrationLinwei Tao, Minjing Dong, Chang XuICML 2023 · 56 citations
- Multimodal Distillation for Egocentric Action RecognitionGorjan Radevski, Dusan Grujicic, Matthew B. Blaschko, Marie-Francine Moens et al.ICCV 2023 · 40 citations
- Jaccard Metric Losses: Optimizing the Jaccard Index with Soft LabelsZifu Wang, Xuefei Ning, Matthew B. BlaschkoNeurIPS 2023 · 39 citations
- A Large-Scale Study of Probabilistic Calibration in Neural Network RegressionVictor Dheur, Souhaib Ben TaiebICML 2023 · 29 citations
Builds on5
- Calibrating Deep Neural Networks using Focal LossJishnu Mukhoti, Viveka Kulharia, Amartya Sanyal, Stuart Golodetz et al.NeurIPS 2020 · 674 citations
- Pitfalls of In-Domain Uncertainty Estimation and Ensembling in Deep LearningArsenii Ashukha, Alexander Lyzhov, Dmitry Molchanov, Dmitry P. VetrovICLR 2020 · 354 citations
- Calibration of Neural Networks using SplinesKartik Gupta, Amir Rahimi, Thalaiyasingam Ajanthan, Thomas Mensink et al.ICLR 2021 · 128 citations
- nuScenes: A Multimodal Dataset for Autonomous DrivingHolger Caesar, Varun Bankiti, Alex H. Lang, Sourabh Vora et al.CVPR 2020
- Scalability in Perception for Autonomous Driving: Waymo Open DatasetPei Sun, Henrik Kretzschmar, Xerxes Dotiwalla, Aurelien Chouard et al.CVPR 2020
Related papers
- Calibrated and Sharp Uncertainties in Deep Learning via Density EstimationVolodymyr Kuleshov, Shachi DeshpandeICML 2022 · 44 citations
- Taking a Step Back with KCal: Multi-Class Kernel-Based Calibration for Deep Neural NetworksZhen Lin, Shubhendu Trivedi, Jimeng SunICLR 2023 · 2 citations
- Top-label calibration and multiclass-to-binary reductionsChirag Gupta, Aaditya RamdasICLR 2022 · 51 citations
- Amortized Conditional Normalized Maximum Likelihood: Reliable Out of Distribution Uncertainty EstimationAurick Zhou, Sergey LevineICML 2021 · 5 citations
- Calibrated Reliable Regression using Maximum Mean DiscrepancyPeng Cui, Wenbo Hu, Jun ZhuNeurIPS 2020 · 71 citations
