Reliable Decisions with Threshold Calibration
Roshni Sahoo, Shengjia Zhao, Alyssa Chen, Stefano Ermon
摘要
Decision makers rely on probabilistic forecasts to predict the loss of different decision rules before deployment. When the forecasted probabilities match the true frequencies, predicted losses will be accurate. Although perfect forecasts are typically impossible, probabilities can be calibrated to match the true frequencies on average. However, we find that this average notion of calibration, which is typically used in practice, does not necessarily guarantee accurate decision loss prediction. Specifically in the regression setting, the loss of threshold decisions, which are decisions based on whether the forecasted outcome falls above or below a cutoff, might not be predicted accurately. We propose a stronger notion of calibration called threshold calibration, which is exactly the condition required to ensure that decision loss is predicted accurately for threshold decisions. We provide an efficient algorithm which takes an uncalibrated forecaster as input and provably outputs a threshold-calibrated forecaster. Our procedure allows downstream decision makers to confidently estimate the loss of any threshold decision under any threshold loss function. Empirically, threshold calibration improves decision loss prediction without compromising on the quality of the decisions in two real-world settings: hospital scheduling decisions and resource allocation decisions.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper8
- Human-Aligned Calibration for AI-Assisted Decision MakingNina Corvelo Benz, Manuel Gomez RodriguezNeurIPS 2023 · 被引用 45 次
- Calibrating Multimodal LearningHuan Ma, Qingyang Zhang, Changqing Zhang, Bingzhe Wu 等ICML 2023 · 被引用 42 次
- Calibration by Distribution Matching: Trainable Kernel Calibration MetricsCharlie Marx, Sofian Zalouk, Stefano ErmonNeurIPS 2023 · 被引用 21 次
- Improving Screening Processes via Calibrated Subset SelectionLequn Wang, Thorsten Joachims, Manuel Gomez RodriguezICML 2022 · 被引用 21 次
- On the Within-Group Fairness of Screening ClassifiersNastaran Okati, Stratis Tsirtsis, Manuel Gomez RodriguezICML 2023 · 被引用 4 次
它引用的顶会 Paper5
- Deep Evidential RegressionAlexander Amini, Wilko Schwarting, Ava Soleimany, Daniela RusNeurIPS 2020 · 被引用 777 次
- Performative PredictionJuan C. Perdomo, Tijana Zrnic, Celestine Mendler-Dünner, Moritz HardtICML 2020 · 被引用 422 次
- Improving model calibration with accuracy versus uncertainty optimizationRanganath Krishnan, Omesh TickooNeurIPS 2020 · 被引用 217 次
- Calibrated Reliable Regression using Maximum Mean DiscrepancyPeng Cui, Wenbo Hu, Jun ZhuNeurIPS 2020 · 被引用 71 次
- Individual Calibration with Randomized ForecastingShengjia Zhao, Tengyu Ma, Stefano ErmonICML 2020 · 被引用 69 次
相关 Paper
- Robust Decision-Making with Partially Calibrated ForecastersShayan Kiyani, Hamed Hassani, George J. Pappas, Aaron RothICLR 2026 · 被引用 1 次
- Efficient Calibration for Decision MakingParikshit Gopalan, Konstantinos Stavropoulos, Kunal Talwar, Pranay TankalaSTOC 2026 · 被引用 3 次
- Reconciling Model Multiplicity for Downstream Decision MakingAlly Yalei Du, Dung Daniel T. Ngo, Zhiwei Steven WuICLR 2025
- Predict to Minimize Swap Regret for All Payoff-Bounded TasksLunjia Hu, Yifan WuFOCS 2024 · 被引用 1 次
- Calibrating Predictions to Decisions: A Novel Approach to Multi-Class CalibrationShengjia Zhao, Michael P. Kim, Roshni Sahoo, Tengyu Ma 等NeurIPS 2021 · 被引用 96 次
