Out-of-distribution Generalization in the Presence of Nuisance-Induced Spurious Correlations
Aahlad Manas Puli, Lily H. Zhang, Eric Karl Oermann, Rajesh Ranganath
Abstract
In many prediction problems, spurious correlations are induced by a changing relationship between the label and a nuisance variable that is also correlated with the covariates. For example, in classifying animals in natural images, the background, which is a nuisance, can predict the type of animal. This nuisance-label relationship does not always hold, and the performance of a model trained under one such relationship may be poor on data with a different nuisance-label relationship. To build predictive models that perform well regardless of the nuisance-label relationship, we develop Nuisance-Randomized Distillation (NURD). We introduce the nuisance-randomized distribution, a distribution where the nuisance and the label are independent. Under this distribution, we define the set of representations such that conditioning on any member, the nuisance and the label remain independent. We prove that the representations in this set always perform better than chance, while representations outside of this set may not. NURD finds a representation from this set that is most informative of the label under the nuisance-randomized distribution, and we prove that this representation achieves the highest performance regardless of the nuisance-label relationship. We evaluate NURD on several tasks including chest X-ray classification where, using non-lung patches as the nuisance, NURD produces models that predict pneumonia under strong spurious correlations.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext b930793b-2d66-45f0-b748-a685657fab91Cited by top-tier papers17
- Change is Hard: A Closer Look at Subpopulation ShiftYuzhe Yang, Haoran Zhang, Dina Katabi, Marzyeh GhassemiICML 2023 · 149 citations
- Don't blame Dataset Shift! Shortcut Learning due to Gradients and Cross EntropyAahlad Manas Puli, Lily H. Zhang, Yoav Wald, Rajesh RanganathNeurIPS 2023 · 37 citations
- Spurious Feature Diversification Improves Out-of-distribution GeneralizationYong Lin, Lu Tan, Yifan Hao, Honam Wong et al.ICLR 2024 · 34 citations
- Spuriosity Didn't Kill the Classifier: Using Invariant Predictions to Harness Spurious FeaturesCian Eastwood, Shashank Singh, Andrei Liviu Nicolicioiu, Marin Vlastelica Pogancic et al.NeurIPS 2023 · 29 citations
- Are All Spurious Features in Natural Language Alike? An Analysis through a Causal LensNitish Joshi, Xiang Pan, He HeEMNLP 2022 · 19 citations
Builds on12
- Out-of-Distribution Generalization via Risk Extrapolation (REx)David Krueger, Ethan Caballero, Jörn-Henrik Jacobsen, Amy Zhang et al.ICML 2021 · 1,163 citations
- Environment Inference for Invariant LearningElliot Creager, Jörn-Henrik Jacobsen, Richard S. ZemelICML 2021 · 454 citations
- An Investigation of Why Overparameterization Exacerbates Spurious CorrelationsShiori Sagawa, Aditi Raghunathan, Pang Wei Koh, Percy LiangICML 2020 · 436 citations
- Fairness without Demographics through Adversarially Reweighted LearningPreethi Lahoti, Alex Beutel, Jilin Chen, Kang Lee et al.NeurIPS 2020 · 406 citations
- Domain Generalization using Causal MatchingDivyat Mahajan, Shruti Tople, Amit SharmaICML 2021 · 399 citations
Related papers
- Discovering and Overcoming Limitations of Noise-engineered Data-free Knowledge DistillationPiyush Raikwar, Deepak MishraNeurIPS 2022 · 22 citations
- Spuriousness-Aware Meta-Learning for Learning Robust ClassifiersGuangtao Zheng, Wenqian Ye, Aidong ZhangKDD 2024 · 3 citations
- ELDET: Early-Learning Distillation with Noisy Labels for Object DetectionDongmin Choi, Sangbin Lee, EungGu Yun, Jonghyuk Baek et al.NeurIPS 2025 · 1 citation
- Novelty Detection Via BlurringSung-Ik Choi, Sae-Young ChungICLR 2020 · 37 citations
- Masked Images Are Counterfactual Samples for Robust Fine-TuningYao Xiao, Ziyi Tang, Pengxu Wei, Cong Liu et al.CVPR 2023
