Invariant Causal Representation Learning for Out-of-Distribution Generalization
Chaochao Lu, Yuhuai Wu, José Miguel Hernández-Lobato, Bernhard Schölkopf
摘要
Due to spurious correlations, machine learning systems often fail to generalize to environments whose distributions differ from the ones used at training time. Prior work addressing this, either explicitly or implicitly, attempted to find a data representation that has an invariant relationship with the target. This is done by leveraging a diverse set of training environments to reduce the effect of spurious features and build an invariant predictor. However, these methods have generalization guarantees only when both data representation and classifiers come from a linear model class. We propose invariant Causal Representation Learning (iCaRL), an approach that enables out-of-distribution (OOD) generalization in the nonlinear setting (i.e., nonlinear representations and nonlinear classifiers). It builds upon a practical and general assumption: the prior over the data representation (i.e., a set of latent variables encoding the data) given the target and the environment belongs to general exponential family distributions, i.e., a more flexible conditionally non-factorized prior that can actually capture complicated dependences between the latent variables. Based on this, we show that it is possible to identify the data representation up to simple transformations. We also show that all direct causes of the target can be fully discovered, which further enables us to obtain generalization guarantees in the nonlinear setting. Experiments on both synthetic and real-world datasets demonstrate that our approach outperforms a variety of baseline methods.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
引用它的顶会 Paper41
- Nonparametric Identifiability of Causal Representations from Unknown InterventionsJulius von Kügelgen, Michel Besserve, Wendong Liang, Luigi Gresele 等NeurIPS 2023 · 被引用 127 次
- Unleashing the Power of Graph Data Augmentation on Covariate Distribution ShiftYongduo Sui, Qitian Wu, Jiancan Wu, Qing Cui 等NeurIPS 2023 · 被引用 63 次
- Joint Learning of Label and Environment Causal Independence for Graph Out-of-Distribution GeneralizationShurui Gui, Meng Liu, Xiner Li, Youzhi Luo 等NeurIPS 2023 · 被引用 54 次
- Leveraging sparse and shared feature activations for disentangled representation learningMarco Fumero, Florian Wenzel, Luca Zancato, Alessandro Achille 等NeurIPS 2023 · 被引用 42 次
- Improving Non-Transferable Representation Learning by Harnessing Content and StyleZiming Hong, Zhenyi Wang, Li Shen, Yu Yao 等ICLR 2024 · 被引用 37 次
相关 Paper
- Identifying Representations for Intervention ExtrapolationSorawit Saengkyongam, Elan Rosenfeld, Pradeep Kumar Ravikumar, Niklas Pfister 等ICLR 2024 · 被引用 20 次
- The Risks of Invariant Risk MinimizationElan Rosenfeld, Pradeep Kumar Ravikumar, Andrej RisteskiICLR 2021 · 被引用 356 次
- Diagnosing and Rectifying Fake OOD Invariance: A Restructured Causal ApproachZiliang Chen, Yongsen Zheng, Zhao-Rong Lai, Quanlong Guan 等AAAI 2024 · 被引用 4 次
- Out-of-distribution Generalization with Causal Invariant TransformationsRuoyu Wang, Mingyang Yi, Zhitang Chen, Shengyu ZhuCVPR 2022 · 被引用 40 次
- Causal Transportability for Visual RecognitionChengzhi Mao, Kevin Xia, James Wang, Hao Wang 等CVPR 2022 · 被引用 27 次
