Obtaining Calibrated Probabilities with Personalized Ranking Models
Wonbin Kweon, SeongKu Kang, Hwanjo Yu
Abstract
For personalized ranking models, the well-calibrated probability of an item being preferred by a user has great practical value. While existing work shows promising results in image classification, probability calibration has not been much explored for personalized ranking. In this paper, we aim to estimate the calibrated probability of how likely a user will prefer an item. We investigate various parametric distributions and propose two parametric calibration methods, namely Gaussian calibration and Gamma calibration. Each proposed method can be seen as a post-processing function that maps the ranking scores of pre-trained models to well-calibrated preference probabilities, without affecting the recommendation performance. We also design the unbiased empirical risk minimization framework that guides the calibration methods to learning of true preference probability from the biased user-item interaction dataset. Extensive evaluations with various personalized ranking models on real-world datasets show that both the proposed calibration methods and the unbiased empirical risk minimization significantly improve the calibration performance.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 489c255b-45b1-4c8d-bb4c-eaa732e97ecaCited by top-tier papers9
- Doubly Calibrated Estimator for Recommendation on Data Missing Not at RandomWonbin Kweon, Hwanjo YuWWW 2024 · 23 citations
- Uncertainty Quantification and Decomposition for LLM-based RecommendationWonbin Kweon, Sanghwan Jang, SeongKu Kang, Hwanjo YuWWW 2025 · 13 citations
- Stability and Multigroup Fairness in Ranking with Uncertain PredictionsSiddartha Devic, Aleksandra Korolova, David Kempe, Vatsal SharanICML 2024 · 9 citations
- Unconstrained Monotonic Calibration of Predictions in Deep Ranking SystemsYimeng Bai, Shunyu Zhang, Yang Zhang, Hu Liu et al.SIGIR 2025 · 2 citations
- MCNet: Monotonic Calibration Networks for Expressive Uncertainty Calibration in Online AdvertisingQuanyu Dai, Jiaren Xiao, Zhaocheng Du, Jieming Zhu et al.WWW 2025 · 2 citations
Builds on3
- LightGCN: Simplifying and Powering Graph Convolution Network for RecommendationXiangnan He, Kuan Deng, Xiang Wang, Yan Li et al.SIGIR 2020 · 4,448 citations
- Calibrating Deep Neural Networks using Focal LossJishnu Mukhoti, Viveka Kulharia, Amartya Sanyal, Stuart Golodetz et al.NeurIPS 2020 · 674 citations
- Intra Order-preserving Functions for Calibration of Multi-Class Neural NetworksAmir Rahimi, Amirreza Shaban, Ching-An Cheng, Richard Hartley et al.NeurIPS 2020 · 96 citations
Related papers
- PAC-Bayes Analysis for Recalibration in ClassificationMasahiro Fujisawa, Futoshi FutamiICML 2025
- Measuring and Mitigating Item Under-Recommendation Bias in Personalized Ranking SystemsZiwei Zhu, Jianling Wang, James CaverleeSIGIR 2020 · 103 citations
- Calibrated Preference Learning: The Case of Label RankingSanto Thies, Viktor Bengs, Timo Kaufmann, Sebastian Vollmer et al.ICML 2026
- Calibration by Distribution Matching: Trainable Kernel Calibration MetricsCharlie Marx, Sofian Zalouk, Stefano ErmonNeurIPS 2023 · 21 citations
- Top-Personalized-K RecommendationWonbin Kweon, SeongKu Kang, Sanghwan Jang, Hwanjo YuWWW 2024
