MBCT: Tree-Based Feature-Aware Binning for Individual Uncertainty Calibration
Siguang Huang, Yunli Wang, Lili Mou, Huayue Zhang, Han Zhu, Chuan Yu, Bo Zheng
摘要
Most machine learning classifiers only concern classification accuracy, while certain applications (such as medical diagnosis, meteorological forecasting, and computation advertising) require the model to predict the true probability, known as a calibrated estimate. In previous work, researchers have developed several calibration methods to post-process the outputs of a predictor to obtain calibrated values, such as binning and scaling methods. Compared with scaling, binning methods are shown to have distribution-free theoretical guarantees, which motivates us to prefer binning methods for calibration. However, we notice that existing binning methods have several drawbacks: (a) the binning scheme only considers the original prediction values, thus limiting the calibration performance; and (b) the binning approach is non-individual, mapping multiple samples in a bin to the same value, and thus is not suitable for order-sensitive applications. In this paper, we propose a featureaware binning framework, called Multiple Boosting Calibration Trees (MBCT), along with a multi-view calibration loss to tackle the above issues. Our MBCT optimizes the binning scheme by the tree structures of features, and adopts a linear function in a tree node to achieve individual calibration. Our MBCT is non-monotonic, and has the potential to improve order accuracy, due to its learnable binning scheme and the individual calibration. We conduct comprehensive experiments on three datasets in different fields. Results show that our method outperforms all competing models in terms of both calibration error and order accuracy. We also conduct simulation experiments, justifying that the proposed multi-view calibration loss is a better metric in modeling calibration error. In addition, our approach is deployed in a real-world online advertising platform; an A/B test over two weeks further demonstrates the effectiveness and great business value of our approach. CCS CONCEPTS • Computing methodologies → Machine learning.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- Adaptive Neural Ranking Framework: Toward Maximized Business Goal for Cascade Ranking SystemsYunli Wang, Zhiqiang Wang, Jian Yang, Shiyang Wen 等WWW 2024 · 被引用 16 次
- Unconstrained Monotonic Calibration of Predictions in Deep Ranking SystemsYimeng Bai, Shunyu Zhang, Yang Zhang, Hu Liu 等SIGIR 2025 · 被引用 2 次
- MCNet: Monotonic Calibration Networks for Expressive Uncertainty Calibration in Online AdvertisingQuanyu Dai, Jiaren Xiao, Zhaocheng Du, Jieming Zhu 等WWW 2025 · 被引用 2 次
- Learning Cascade Ranking as One NetworkYunli Wang, Zhen Zhang, Zhiqiang Wang, Zixuan Yang 等ICML 2025
- GETS: Ensemble Temperature Scaling for Calibration in Graph Neural NetworksDingyi Zhuang, Chonghe Jiang, Yunhan Zheng, Shenhao Wang 等ICLR 2025
它引用的顶会 Paper3
- Mix-n-Match : Ensemble and Compositional Methods for Uncertainty Calibration in Deep LearningJize Zhang, Bhavya Kailkhura, Thomas Yong-Jin HanICML 2020 · 被引用 276 次
- Distribution-free binary classification: prediction sets, confidence intervals and calibrationChirag Gupta, Aleksandr Podkopaev, Aaditya RamdasNeurIPS 2020 · 被引用 105 次
- Individual Calibration with Randomized ForecastingShengjia Zhao, Tengyu Ma, Stefano ErmonICML 2020 · 被引用 69 次
相关 Paper
- PAC-Bayes Analysis for Recalibration in ClassificationMasahiro Fujisawa, Futoshi FutamiICML 2025
- Discretization-free Multicalibration through Loss Minimization over Tree EnsemblesHongyi Henry Jin, Zijun Ding, Dung Daniel T. Ngo, Zhiwei Steven WuNeurIPS 2025 · 被引用 6 次
- Top-label calibration and multiclass-to-binary reductionsChirag Gupta, Aaditya RamdasICLR 2022 · 被引用 51 次
- Field-aware Calibration: A Simple and Empirically Strong Method for Reliable Probabilistic PredictionsFeiyang Pan, Xiang Ao, Pingzhong Tang, Min Lu 等WWW 2020 · 被引用 30 次
- Calibration tests beyond classificationDavid Widmann, Fredrik Lindsten, Dave ZachariahICLR 2021 · 被引用 23 次
