When Personalization Harms Performance: Reconsidering the Use of Group Attributes in Prediction
Vinith Menon Suriyakumar, Marzyeh Ghassemi, Berk Ustun
Abstract
Machine learning models are often personalized with categorical attributes that are protected, sensitive, self-reported, or costly to acquire. In this work, we show models that are personalized with group attributes can reduce performance at a group level. We propose formal conditions to ensure the"fair use"of group attributes in prediction tasks by training one additional model -- i.e., collective preference guarantees to ensure that each group who provides personal data will receive a tailored gain in performance in return. We present sufficient conditions to ensure fair use in empirical risk minimization and characterize failure modes that lead to fair use violations due to standard practices in model development and deployment. We present a comprehensive empirical study of fair use in clinical prediction tasks. Our results demonstrate the prevalence of fair use violations in practice and illustrate simple interventions to mitigate their harm.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 489964c9-823b-49ec-8ec5-441a91c16529Cited by top-tier papers2
- "Who experiences large model decay and why?" A Hierarchical Framework for Diagnosing Heterogeneous Performance DriftHarvineet Singh, Fan Xia, Alexej Gossmann, Andrew Chuang et al.ICML 2025
- Regretful Decisions under Label NoiseSujay Nagaraj, Yang Liu, Flávio P. Calmon, Berk UstunICLR 2025
Builds on5
- Minimax Pareto Fairness: A Multi Objective PerspectiveNatalia Martínez, Martín Bertrán, Guillermo SapiroICML 2020 · 232 citations
- Online Certification of Preference-Based Fairness for Personalized Recommender SystemsVirginie Do, Sam Corbett-Davies, Jamal Atif, Nicolas UsunierAAAI 2022 · 47 citations
- Model Distillation for Revenue Optimization: Interpretable Personalized PricingMax Biggs, Wei Sun, Markus EttlICML 2021 · 42 citations
- Learning Optimal Predictive ChecklistsHaoran Zhang, Quaid Morris, Berk Ustun, Marzyeh GhassemiNeurIPS 2021 · 15 citations
- On the Epistemic Limits of Personalized PredictionLucas Monteiro Paes, Carol Xuan Long, Berk Ustun, Flávio P. CalmonNeurIPS 2022 · 14 citations
Related papers
- Participatory Personalization in ClassificationHailey Joren, Chirag Nagpal, Katherine A. Heller, Berk UstunNeurIPS 2023 · 7 citations
- Fair Learning with Private Demographic DataHussein Mozannar, Mesrob I. Ohannessian, Nathan SrebroICML 2020 · 85 citations
- When do Minimax-fair Learning and Empirical Risk Minimization Coincide?Harvineet Singh, Matthäus Kleindessner, Volkan Cevher, Rumi Chunara et al.ICML 2023 · 6 citations
- Individual Arbitrariness and Group FairnessCarol Xuan Long, Hsiang Hsu, Wael Alghamdi, Flávio P. CalmonNeurIPS 2023 · 16 citations
- Multiaccuracy and Multicalibration via Proxy GroupsBeepul Bharti, Mary Versa Clemens-Sewall, Paul H. Yi, Jeremias SulamICML 2025
