Causally motivated multi-shortcut identification and removal
Jiayun Zheng, Maggie Makar
摘要
For predictive models to provide reliable guidance in decision making processes, they are often required to be accurate and robust to distribution shifts. Shortcut learning-where a model relies on spurious correlations or shortcuts to predict the target label-undermines the robustness property, leading to models with poor out-of-distribution accuracy despite good in-distribution performance. Existing work on shortcut learning either assumes that the set of possible shortcuts is known a priori or is discoverable using interpretability methods such as saliency maps, which might not always be true. Instead, we propose a two step approach to (1) efficiently identify relevant shortcuts, and (2) leverage the identified shortcuts to build models that are robust to distribution shifts. Our approach relies on having access to a (possibly) high dimensional set of auxiliary labels at training time, some of which correspond to possible shortcuts. We show both theoretically and empirically that our approach is able to identify a sufficient set of shortcuts leading to more efficient predictors in finite samples.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- Beyond Invariance: Test-Time Label-Shift Adaptation for Addressing "Spurious" CorrelationsQingyao Sun, Kevin P. Murphy, Sayna Ebrahimi, Alexander D'AmourNeurIPS 2023 · 被引用 10 次
- Causal Effect Regularization: Automated Detection and Removal of Spurious CorrelationsAbhinav Kumar, Amit Deshpande, Amit SharmaNeurIPS 2023 · 被引用 7 次
- Debugging Concept Bottleneck Models through Removal and RetrainingEric Enouen, Sainyam GalhotraICLR 2026 · 被引用 2 次
- Factored Causal Representation Learning for Robust Reward Modeling in RLHFYupei Yang, Lin Yang, Wanxi Deng, Lin Qu 等ICML 2026 · 被引用 1 次
- Do ImageNet-trained Models Learn Shortcuts? The Impact of Frequency Shortcuts on GeneralizationShunxin Wang, Raymond N. J. Veldhuis, Nicola StrisciuglioCVPR 2025
它引用的顶会 Paper4
- Out-of-Distribution Generalization via Risk Extrapolation (REx)David Krueger, Ethan Caballero, Jörn-Henrik Jacobsen, Amy Zhang 等ICML 2021 · 被引用 1,163 次
- Invariant Causal Representation Learning for Out-of-Distribution GeneralizationChaochao Lu, Yuhuai Wu, José Miguel Hernández-Lobato, Bernhard SchölkopfICLR 2022 · 被引用 119 次
- Counterfactual Invariance to Spurious Correlations in Text ClassificationVictor Veitch, Alexander D'Amour, Steve Yadlowsky, Jacob EisensteinNeurIPS 2021 · 被引用 108 次
- Permutation WeightingDavid Arbour, Drew Dimmery, Arjun SondhiICML 2021 · 被引用 24 次
相关 Paper
- Improving the robustness of NLI models with minimax trainingMichalis Korakakis, Andreas VlachosACL 2023 · 被引用 4 次
- Learning Concept Credible Models for Mitigating ShortcutsJiaxuan Wang, Sarah Jabbour, Maggie Makar, Michael W. Sjoding 等NeurIPS 2022 · 被引用 8 次
- Roadblocks for Temporarily Disabling Shortcuts and Learning New KnowledgeHongjing Niu, Hanting Li, Feng Zhao, Bin LiNeurIPS 2022 · 被引用 9 次
- Prompting is a Double-Edged Sword: Improving Worst-Group Robustness of Foundation ModelsAmrith Setlur, Saurabh Garg, Virginia Smith, Sergey LevineICML 2024 · 被引用 4 次
- Saliency is a Possible Red Herring When Diagnosing Poor GeneralizationJoseph D. Viviano, Becks Simpson, Francis Dutil, Yoshua Bengio 等ICLR 2021 · 被引用 46 次
