Field-aware Calibration: A Simple and Empirically Strong Method for Reliable Probabilistic Predictions
Feiyang Pan, Xiang Ao, Pingzhong Tang, Min Lu, Dapeng Liu, Lei Xiao, Qing He
Abstract
It is often observed that the probabilistic predictions given by a machine learning model can disagree with averaged actual outcomes on specific subsets of data, which is also known as the issue of miscalibration. It is responsible for the unreliability of practical machine learning systems. For example, in online advertising, an ad can receive a click-through rate prediction of 0.1 over some population of users where its actual click rate is 0.15. In such cases, the probabilistic predictions have to be fixed before the system can be deployed. In this paper, we first introduce a new evaluation metric named field-level calibration error that measures the bias in predictions over the sensitive input field that the decision-maker concerns. We show that existing post-hoc calibration methods have limited improvements in the new field-level metric and other non-calibration metrics such as the AUC score. To this end, we propose Neural Calibration, a simple yet powerful post-hoc calibration method that learns to calibrate by making full use of the field-aware information over the validation set. We present extensive experiments on five large-scale datasets. The results showed that Neural Calibration significantly improves against uncalibrated predictions in common metrics such as the negative log-likelihood, Brier score and AUC, as well as the proposed field-level calibration error. CCS CONCEPTS • Computing methodologies → Machine learning.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 34dd7609-1a88-4643-a73f-64342f9eef3aCited by top-tier papers4
- Variable-Based Calibration for Machine Learning ClassifiersMarkelle Kelly, Padhraic SmythAAAI 2023 · 7 citations
- Unconstrained Monotonic Calibration of Predictions in Deep Ranking SystemsYimeng Bai, Shunyu Zhang, Yang Zhang, Hu Liu et al.SIGIR 2025 · 2 citations
- MCNet: Monotonic Calibration Networks for Expressive Uncertainty Calibration in Online AdvertisingQuanyu Dai, Jiaren Xiao, Zhaocheng Du, Jieming Zhu et al.WWW 2025 · 2 citations
- Auto-bidding under Return-on-Spend Constraints with Uncertainty QuantificationJiale Han, Chun Gan, Chengcheng Zhang, Jie He et al.WWW 2026
Related papers
- Calibration tests beyond classificationDavid Widmann, Fredrik Lindsten, Dave ZachariahICLR 2021 · 23 citations
- Calibration Matters: Tackling Maximization Bias in Large-scale Advertising Recommendation SystemsYewen Fan, Nian Si, Kun ZhangICLR 2023
- Meta-Cal: Well-controlled Post-hoc Calibration by RankingXingchen Ma, Matthew B. BlaschkoICML 2021 · 44 citations
- A Large-Scale Study of Probabilistic Calibration in Neural Network RegressionVictor Dheur, Souhaib Ben TaiebICML 2023 · 29 citations
- Taking a Step Back with KCal: Multi-Class Kernel-Based Calibration for Deep Neural NetworksZhen Lin, Shubhendu Trivedi, Jimeng SunICLR 2023 · 2 citations
