When Does Group Invariant Learning Survive Spurious Correlations?
Yimeng Chen, Ruibin Xiong, Zhi-Ming Ma, Yanyan Lan
摘要
By inferring latent groups in the training data, recent works introduce invariant learning to the case where environment annotations are unavailable. Typically, learning group invariance under a majority/minority split is empirically shown to be effective in improving out-of-distribution generalization on many datasets. However, theoretical guarantee for these methods on learning invariant mechanisms is lacking. In this paper, we reveal the insufficiency of existing group invariant learning methods in preventing classifiers from depending on spurious correlations in the training set. Specifically, we propose two criteria on judging such sufficiency. Theoretically and empirically, we show that existing methods can violate both criteria and thus fail in generalizing to spurious correlation shifts. Motivated by this, we design a new group invariant learning method, which constructs groups with statistical independence tests, and reweights samples by group label proportion to meet the criteria. Experiments on both synthetic and real data demonstrate that the new method significantly outperforms existing group invariant learning methods in generalizing to spurious correlation shifts.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper14
- Joint Learning of Label and Environment Causal Independence for Graph Out-of-Distribution GeneralizationShurui Gui, Meng Liu, Xiner Li, Youzhi Luo 等NeurIPS 2023 · 被引用 54 次
- Group Robust Classification Without Any Group InformationChristos Tsirigotis, João Monteiro, Pau Rodríguez, David Vázquez 等NeurIPS 2023 · 被引用 34 次
- Multimodality Invariant Learning for Multimedia-Based New Item RecommendationHaoyue Bai, Le Wu, Min Hou, Miaomiao Cai 等SIGIR 2024 · 被引用 30 次
- Improving Generalization of Alignment with Human Preferences through Group Invariant LearningRui Zheng, Wei Shen, Yuan Hua, Wenbin Lai 等ICLR 2024 · 被引用 25 次
- Explore and Exploit the Diverse Knowledge in Model Zoo for Domain GeneralizationYimeng Chen, Tianyang Hu, Fengwei Zhou, Zhenguo Li 等ICML 2023 · 被引用 14 次
它引用的顶会 Paper19
- In Search of Lost Domain GeneralizationIshaan Gulrajani, David Lopez-PazICLR 2021 · 被引用 1,416 次
- Out-of-Distribution Generalization via Risk Extrapolation (REx)David Krueger, Ethan Caballero, Jörn-Henrik Jacobsen, Amy Zhang 等ICML 2021 · 被引用 1,163 次
- Just Train Twice: Improving Group Robustness without Training Group InformationEvan Zheran Liu, Behzad Haghgoo, Annie S. Chen, Aditi Raghunathan 等ICML 2021 · 被引用 683 次
- Environment Inference for Invariant LearningElliot Creager, Jörn-Henrik Jacobsen, Richard S. ZemelICML 2021 · 被引用 454 次
- Learning from Failure: De-biasing Classifier from Biased ClassifierJun Hyun Nam, Hyuntak Cha, Sungsoo Ahn, Jaeho Lee 等NeurIPS 2020 · 被引用 428 次
相关 Paper
- Heterogeneous Risk MinimizationJiashuo Liu, Zheyuan Hu, Peng Cui, Bo Li 等ICML 2021 · 被引用 170 次
- Improving Group Robustness on Spurious Correlation Requires Preciser Group InferenceYujin Han, Difan ZouICML 2024 · 被引用 13 次
- Mitigating Spurious Correlations in Text Classification Using Latent Space GeometryJiasen Gao, Xiaoliang Chen, Duoqian Miao, Xu Gu 等ACL 2026
- Breaking Correlation Shift via Conditional Invariant RegularizerMingyang Yi, Ruoyu Wang, Jiacheng Sun, Zhenguo Li 等ICLR 2023
- Let Samples Speak: Mitigating Spurious Correlation by Exploiting the Clusterness of SamplesWeiwei Li, Junzhuo Liu, Yuanyuan Ren, Yuchen Zheng 等CVPR 2025
