Predictive Multiplicity in Probabilistic Classification
Jamelle Watson-Daniels, David C. Parkes, Berk Ustun
摘要
Machine learning models are often used to inform real world risk assessment tasks: predicting consumer default risk, predicting whether a person suffers from a serious illness, or predicting a person's risk to appear in court. Given multiple models that perform almost equally well for a prediction task, to what extent do predictions vary across these models? If predictions are relatively consistent for similar models, then the standard approach of choosing the model that optimizes a penalized loss suffices. But what if predictions vary significantly for similar models? In machine learning, this is referred to as predictive multiplicity i.e. the prevalence of conflicting predictions assigned by near-optimal competing models. In this paper, we present a framework for measuring predictive multiplicity in probabilistic classification (predicting the probability of a positive outcome). We introduce measures that capture the variation in risk estimates over the set of competing models, and develop optimization-based methods to compute these measures efficiently and reliably for convex empirical risk minimization problems. We demonstrate the incidence and prevalence of predictive multiplicity in real-world tasks. Further, we provide insight into how predictive multiplicity arises by analyzing the relationship between predictive multiplicity and data set characteristics (outliers, separability, and majority-minority structure). Our results emphasize the need to report predictive multiplicity more widely.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper16
- Robust Counterfactual Explanations for Neural Networks With Probabilistic GuaranteesFaisal Hamman, Erfaun Noorani, Saumitra Mishra, Daniele Magazzeni 等ICML 2023 · 被引用 54 次
- A Path to Simpler Models Starts With NoiseLesia Semenova, Harry Chen, Ronald Parr, Cynthia RudinNeurIPS 2023 · 被引用 41 次
- Exploring the limits of strong membership inference attacks on large language modelsJamie Hayes, Ilia Shumailov, Christopher A. Choquette-Choo, Matthew Jagielski 等NeurIPS 2025 · 被引用 26 次
- Dropout-Based Rashomon Set Exploration for Efficient Predictive Multiplicity EstimationHsiang Hsu, Guihong Li, Shaohan Hu, Chun-Fu ChenICLR 2024 · 被引用 18 次
- Individual Arbitrariness and Group FairnessCarol Xuan Long, Hsiang Hsu, Wael Alghamdi, Flávio P. CalmonNeurIPS 2023 · 被引用 16 次
它引用的顶会 Paper6
- Consistent Estimators for Learning to Defer to an ExpertHussein Mozannar, David A. SontagICML 2020 · 被引用 267 次
- Predictive Multiplicity in ClassificationCharles T. Marx, Flávio P. Calmon, Berk UstunICML 2020 · 被引用 197 次
- Visual Reasoning Strategies for Effect Size Judgments and DecisionsAlex Kale, Matthew Kay, Jessica HullmanIEEE VIS 2020 · 被引用 112 次
- Characterizing Fairness Over the Set of Good Models Under Selective LabelsAmanda Coston, Ashesh Rambachan, Alexandra ChouldechovaICML 2021 · 被引用 98 次
- Selective Ensembles for Consistent PredictionsEmily Black, Klas Leino, Matt FredriksonICLR 2022 · 被引用 29 次
相关 Paper
- Rashomon Capacity: A Metric for Predictive Multiplicity in ClassificationHsiang Hsu, Flávio P. CalmonNeurIPS 2022 · 被引用 65 次
- Implications of Model Indeterminacy for Explanations of Automated DecisionsMarc-Etienne Brunet, Ashton Anderson, Richard S. ZemelNeurIPS 2022 · 被引用 22 次
- Perceptions of the Fairness Impacts of Multiplicity in Machine LearningAnna P. Meyer, Yea-Seul Kim, Loris D'Antoni, Aws AlbarghouthiCHI 2025 · 被引用 1 次
- When Personalization Harms Performance: Reconsidering the Use of Group Attributes in PredictionVinith Menon Suriyakumar, Marzyeh Ghassemi, Berk UstunICML 2023 · 被引用 10 次
- Reconciling Model Multiplicity for Downstream Decision MakingAlly Yalei Du, Dung Daniel T. Ngo, Zhiwei Steven WuICLR 2025
