When Machine Learning Gets Personal: Evaluating Prediction and Explanation
Louisa Cornelis, Guillermo Bernardez, Haewon Jeong, Nina Miolane
Abstract
In high-stakes domains like healthcare, users often expect that sharing personal information with machine learning systems will yield tangible benefits, such as more accurate diagnoses and clearer explanations of contributing factors. However, the validity of this assumption remains largely unexplored. We propose a unified framework to quantify how personalizing a model influences both prediction and explanation. We show that its impacts on prediction and explanation can diverge: a model may become more or less explainable even when prediction is unchanged. For practical settings, we study a standard hypothesis test for detecting personalization effects on demographic groups. We derive a finite-sample lower bound on its probability of error as a function of group sizes, number of personal attributes, and desired benefit from personalization. This provides actionable insights, such as which dataset characteristics are necessary to test an effect, or the maximum effect that can be tested given a dataset. We apply our framework to real-world tabular datasets using feature-attribution methods, uncovering scenarios where effects are fundamentally untestable due to the dataset statistics. Our results highlight the need for joint evaluation of prediction and explanation in personalized models and the importance of designing models and datasets with sufficient information for such evaluation.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 90b0b928-5ed0-4b80-899e-5ea0daf17cd1Builds on4
- Framework for Evaluating Faithfulness of Local ExplanationsSanjoy Dasgupta, Nave Frost, Michal MoshkovitzICML 2022 · 87 citations
- ERASER: A Benchmark to Evaluate Rationalized NLP ModelsJay DeYoung, Sarthak Jain, Nazneen Fatema Rajani, Eric P. Lehman et al.ACL 2020 · 36 citations
- On the Sensitivity and Stability of Model Interpretations in NLPFan Yin, Zhouxing Shi, Cho-Jui Hsieh, Kai-Wei ChangACL 2022 · 35 citations
- On the Epistemic Limits of Personalized PredictionLucas Monteiro Paes, Carol Xuan Long, Berk Ustun, Flávio P. CalmonNeurIPS 2022 · 14 citations
Related papers
- When Personalization Harms Performance: Reconsidering the Use of Group Attributes in PredictionVinith Menon Suriyakumar, Marzyeh Ghassemi, Berk UstunICML 2023 · 10 citations
- Data Minimization at Inference TimeCuong Tran, Ferdinando FiorettoNeurIPS 2023 · 8 citations
- Participatory Personalization in ClassificationHailey Joren, Chirag Nagpal, Katherine A. Heller, Berk UstunNeurIPS 2023 · 7 citations
- Does Explainable Artificial Intelligence Improve Human Decision-Making?Yasmeen Alufaisan, Laura R. Marusich, Jonathan Z. Bakdash, Yan Zhou et al.AAAI 2021 · 135 citations
- Domain constraints improve risk prediction when outcome data is missingSidhika Balachandar, Nikhil Garg, Emma PiersonICLR 2024 · 11 citations
