A hierarchical decomposition for explaining ML performance discrepancies
Harvineet Singh, Fan Xia, Adarsh Subbaswamy, Alexej Gossmann, Jean Feng
摘要
Machine learning (ML) algorithms can often differ in performance across domains. Understanding their performance differs is crucial for determining what types of interventions (e.g., algorithmic or operational) are most effective at closing the performance gaps. Existing methods focus on of the total performance gap into the impact of a shift in the distribution of features versus the impact of a shift in the conditional distribution of the outcome ; however, such coarse explanations offer only a few options for how one can close the performance gap. that quantify the importance of each variable to each term in the aggregate decomposition can provide a much deeper understanding and suggest much more targeted interventions. However, existing methods assume knowledge of the full causal graph or make strong parametric assumptions. We introduce a nonparametric hierarchical framework that provides both aggregate and detailed decompositions for explaining why the performance of an ML algorithm differs across domains, without requiring causal knowledge. We derive debiased, computationally-efficient estimators, and statistical inference procedures for asymptotically valid confidence intervals.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- "Who experiences large model decay and why?" A Hierarchical Framework for Diagnosing Heterogeneous Performance DriftHarvineet Singh, Fan Xia, Alexej Gossmann, Andrew Chuang 等ICML 2025
- Explaining Concept Shift with Interpretable Feature AttributionRuiqi Lyu, Alistair Turcan, Bryan WilderICML 2026
- Going Beyond Static: Understanding Shifts with Time-Series AttributionJiashuo Liu, Nabeel Seedat, Peng Cui, Mihaela van der SchaarICLR 2025
它引用的顶会 Paper5
- Efficient nonparametric statistical inference on population feature importance using Shapley valuesBrian D. Williamson, Jean FengICML 2020 · 被引用 86 次
- Feature Shift Detection: Localizing Which Features Have Shifted via Conditional Distribution TestsSean Kulinski, Saurabh Bagchi, David I. InouyeNeurIPS 2020 · 被引用 39 次
- Towards Explaining Distribution ShiftsSean Kulinski, David I. InouyeICML 2023 · 被引用 38 次
- "Why did the Model Fail?": Attributing Model Performance Changes to Distribution ShiftsHaoran Zhang, Harvineet Singh, Marzyeh Ghassemi, Shalmali JoshiICML 2023 · 被引用 37 次
- Distilling Model Failures as Directions in Latent SpaceSaachi Jain, Hannah Lawrence, Ankur Moitra, Aleksander MadryICLR 2023 · 被引用 11 次
相关 Paper
- Explaining Algorithmic Fairness Through Fairness-Aware Causal Path DecompositionWeishen Pan, Sen Cui, Jiang Bian, Changshui Zhang 等KDD 2021 · 被引用 29 次
- Understanding challenges to the interpretation of disaggregated evaluations of algorithmic fairnessStephen Pfohl, Natalie Harris, Chirag Nagpal, David Madras 等NeurIPS 2025 · 被引用 9 次
- Debiasing Concept-based Explanations with Causal AnalysisMohammad Taha Bahadori, David HeckermanICLR 2021 · 被引用 8 次
- Returning The Favour: When Regression Benefits From Probabilistic Causal KnowledgeShahine Bouabid, Jake Fawkes, Dino SejdinovicICML 2023
- Interpretations are Useful: Penalizing Explanations to Align Neural Networks with Prior KnowledgeLaura Rieger, Chandan Singh, W. James Murdoch, Bin YuICML 2020 · 被引用 249 次
