"Why did the Model Fail?": Attributing Model Performance Changes to Distribution Shifts
Haoran Zhang, Harvineet Singh, Marzyeh Ghassemi, Shalmali Joshi
摘要
Machine learning models frequently experience performance drops under distribution shifts. The underlying cause of such shifts may be multiple simultaneous factors such as changes in data quality, differences in specific covariate distributions, or changes in the relationship between label and features. When a model does fail during deployment, attributing performance change to these factors is critical for the model developer to identify the root cause and take mitigating actions. In this work, we introduce the problem of attributing performance differences between environments to distribution shifts in the underlying data generating mechanisms. We formulate the problem as a cooperative game where the players are distributions. We define the value of a set of distributions to be the change in model performance when only this set of distributions has changed between environments, and derive an importance weighting method for computing the value of an arbitrary set of distributions. The contribution of each distribution to the total performance change is then quantified as its Shapley value. We demonstrate the correctness and utility of our method on synthetic, semi-synthetic, and real-world case studies, showing its effectiveness in attributing performance changes to a wide range of distribution shifts.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper10
- Self-Healing Machine Learning: A Framework for Autonomous Adaptation in Real-World EnvironmentsPaulius Rauba, Nabeel Seedat, Krzysztof Kacprzyk, Mihaela van der SchaarNeurIPS 2024 · 被引用 15 次
- A hierarchical decomposition for explaining ML performance discrepanciesHarvineet Singh, Fan Xia, Adarsh Subbaswamy, Alexej Gossmann 等NeurIPS 2024 · 被引用 9 次
- EigenScore: OOD Detection using Posterior Covariance in Diffusion ModelsShirin Shoushtari, Yi Wang, Xiao Shi, M. Salman Asif 等ICLR 2026 · 被引用 5 次
- ICYM2I: The illusion of multimodal informativeness under missingnessYoung Sang Choi, Vincent Jeanselme, Pierre A. Elias, Shalmali JoshiICLR 2026 · 被引用 2 次
- Path-specific effects for pulse-oximetry guided decisions in critical careKevin Zhang, Yonghan Jung, Divyat Mahajan, Karthikeyan Shanmugam 等NeurIPS 2025 · 被引用 2 次
它引用的顶会 Paper11
- WILDS: A Benchmark of in-the-Wild Distribution ShiftsPang Wei Koh, Shiori Sagawa, Henrik Marklund, Sang Michael Xie 等ICML 2021 · 被引用 1,773 次
- The Many Shapley Values for Model ExplanationMukund Sundararajan, Amir NajmiICML 2020 · 被引用 799 次
- A Fine-Grained Analysis on Distribution ShiftOlivia Wiles, Sven Gowal, Florian Stimberg, Sylvestre-Alvise Rebuffi 等ICLR 2022 · 被引用 258 次
- Change is Hard: A Closer Look at Subpopulation ShiftYuzhe Yang, Haoran Zhang, Dina Katabi, Marzyeh GhassemiICML 2023 · 被引用 149 次
- Feature Shift Detection: Localizing Which Features Have Shifted via Conditional Distribution TestsSean Kulinski, Saurabh Bagchi, David I. InouyeNeurIPS 2020 · 被引用 39 次
相关 Paper
- Explaining Probabilistic Models with Distributional ValuesLuca Franceschi, Michele Donini, Cédric Archambeau, Matthias W. SeegerICML 2024 · 被引用 4 次
- Multiply-Robust Causal Change AttributionVictor Quintas-Martinez, Mohammad Taha Bahadori, Eduardo Santiago, Jeff Mu 等ICML 2024 · 被引用 5 次
- Problems with Shapley-value-based explanations as feature importance measuresI. Elizabeth Kumar, Suresh Venkatasubramanian, Carlos Scheidegger, Sorelle A. FriedlerICML 2020 · 被引用 458 次
- Rethinking Shapley Value for Negative Interactions in Non-convex GamesWonjoon Chang, Myeongjin Lee, Jaesik ChoiICLR 2025
- Explanatory Model Monitoring to Understand the Effects of Feature Shifts on PerformanceThomas Decker, Alexander Koebler, Michael Lebacher, Ingo Thon 等KDD 2024
