Spuriosity Didn't Kill the Classifier: Using Invariant Predictions to Harness Spurious Features
Cian Eastwood, Shashank Singh, Andrei Liviu Nicolicioiu, Marin Vlastelica Pogancic, Julius von Kügelgen, Bernhard Schölkopf
Abstract
To avoid failures on out-of-distribution data, recent works have sought to extract features that have a stable or invariant relationship with the label across domains, discarding the"spurious"or unstable features whose relationship with the label changes across domains. However, unstable features often carry complementary information about the label that could boost performance if used correctly in the test domain. Our main contribution is to show that it is possible to learn how to use these unstable features in the test domain without labels. In particular, we prove that pseudo-labels based on stable features provide sufficient guidance for doing so, provided that stable and unstable features are conditionally independent given the label. Based on this theoretical insight, we propose Stable Feature Boosting (SFB), an algorithm for: (i) learning a predictor that separates stable and conditionally-independent unstable features; and (ii) using the stable-feature predictions to adapt the unstable-feature predictions in the test domain. Theoretically, we prove that SFB can learn an asymptotically-optimal predictor without test-domain labels. Empirically, we demonstrate the effectiveness of SFB on real and synthetic data.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 3fc55b7b-0b5c-4860-9cd0-e1da49bfea4aCited by top-tier papers8
- Nonparametric Identifiability of Causal Representations from Unknown InterventionsJulius von Kügelgen, Michel Besserve, Wendong Liang, Luigi Gresele et al.NeurIPS 2023 · 127 citations
- Domain Generalisation via Imprecise LearningAnurag Singh, Siu Lun Chau, Shahine Bouabid, Krikamol MuandetICML 2024 · 16 citations
- When Shift Happens - Confounding Is to BlameAbbavaram Gowtham Reddy, Celia Rubio-Madrigal, Rebekka Burkholz, Krikamol MuandetICLR 2026 · 5 citations
- Dual-Path Counterfactual Integration for Multimodal Aspect-Based Sentiment ClassificationRui Liu, Jiahao Cao, Jiaqian Ren, Xu Bai et al.EMNLP 2025 · 1 citation
- ERICT: Enhancing Robustness by Identifying Concept Tokens in Zero-Shot Vision Language ModelsXinpeng Dong, Min Zhang, Didi Zhu, Ye Jun Jian et al.ICML 2025
Builds on20
- WILDS: A Benchmark of in-the-Wild Distribution ShiftsPang Wei Koh, Shiori Sagawa, Henrik Marklund, Sang Michael Xie et al.ICML 2021 · 1,773 citations
- Tent: Fully Test-Time Adaptation by Entropy MinimizationDequan Wang, Evan Shelhamer, Shaoteng Liu, Bruno A. Olshausen et al.ICLR 2021 · 1,731 citations
- Do We Really Need to Access the Source Data? Source Hypothesis Transfer for Unsupervised Domain AdaptationJian Liang, Dapeng Hu, Jiashi FengICML 2020 · 1,624 citations
- Distributionally Robust Neural NetworksShiori Sagawa, Pang Wei Koh, Tatsunori B. Hashimoto, Percy LiangICLR 2020 · 1,578 citations
- In Search of Lost Domain GeneralizationIshaan Gulrajani, David Lopez-PazICLR 2021 · 1,416 citations
Related papers
- Learning Stable Classifiers by Transferring Unstable FeaturesYujia Bao, Shiyu Chang, Regina BarzilayICML 2022 · 8 citations
- Self-training Avoids Using Spurious Features Under Domain ShiftYining Chen, Colin Wei, Ananya Kumar, Tengyu MaNeurIPS 2020 · 100 citations
- An Adaptive Hybrid Framework for Cross-domain Aspect-based Sentiment AnalysisYan Zhou, Fuqing Zhu, Pu Song, Jizhong Han et al.AAAI 2021 · 33 citations
- A Theoretical Analysis on Independence-driven Importance Weighting for Covariate-shift GeneralizationRenzhe Xu, Xingxuan Zhang, Zheyan Shen, Tong Zhang et al.ICML 2022 · 36 citations
- Unsupervised Domain Adaptation via Structured Prediction Based Selective Pseudo-LabelingQian Wang, Toby P. BreckonAAAI 2020 · 257 citations
