Recovering Latent Causal Factor for Generalization to Distributional Shifts
Xinwei Sun, Botong Wu, Xiangyu Zheng, Chang Liu, Wei Chen, Tao Qin, Tie-Yan Liu
摘要
Distributional shifts between training and target domains may degrade the prediction accuracy of learned models, mainly because these models often learn features that possess only correlation rather than causal relation with the output. Such a correlation, which is known as "spurious correlation" statistically, is domaindependent hence may fail to generalize to unseen domains. To avoid such a spurious correlation, we propose Latent Causal Invariance Models (LaCIM) that specifies the underlying causal structure of the data and the source of distributional shifts, guiding us to pursue only causal factor for prediction. Specifically, the LaCIM introduces a pair of correlated latent factors: (a) causal factor and (b) others, while the extent of this correlation is governed by a domain variable that characterizes the distributional shifts. On the basis of this, we prove that the distribution of observed variables conditioning on latent variables is shift-invariant. Equipped with such an invariance, we prove that the causal factor can be recovered without mixing information from others, which induces the ground-truth predicting mechanism. We propose a Variational-Bayesian-based method to learn this invariance for prediction. The utility of our approach is verified by improved generalization to distributional shifts on various real-world data. Our code is freely available at https://github.com/wubotong/LaCIM .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper14
- OoD-Bench: Quantifying and Understanding Two Dimensions of Out-of-Distribution GeneralizationNanyang Ye, Kaican Li, Haoyue Bai, Runpeng Yu 等CVPR 2022 · 被引用 74 次
- DomainDrop: Suppressing Domain-Sensitive Channels for Domain GeneralizationJintao Guo, Lei Qi, Yinghuan ShiICCV 2023 · 被引用 47 次
- Mix and Reason: Reasoning over Semantic Topology with Data Mixing for Domain GeneralizationChaoqi Chen, Luyao Tang, Feng Liu, Gangming Zhao 等NeurIPS 2022 · 被引用 43 次
- An Adaptive Kernel Approach to Federated Learning of Heterogeneous Causal EffectsThanh Vinh Vo, Arnab Bhattacharyya, Young Lee, Tze-Yun LeongNeurIPS 2022 · 被引用 29 次
- CODA: Generalizing to Open and Unseen Domains with Compaction and DisambiguationChaoqi Chen, Luyao Tang, Yue Huang, Xiaoguang Han 等NeurIPS 2023 · 被引用 17 次
它引用的顶会 Paper4
- FaceForensics++: Learning to Detect Manipulated Facial ImagesAndreas Rössler, Davide Cozzolino, Luisa Verdoliva, Christian Riess 等ICCV 2019 · 被引用 2,966 次
- Out-of-Distribution Generalization via Risk Extrapolation (REx)David Krueger, Ethan Caballero, Jörn-Henrik Jacobsen, Amy Zhang 等ICML 2021 · 被引用 1,163 次
- ICE-BeeM: Identifiable Conditional Energy-Based Deep Models Based on Nonlinear ICAIlyes Khemakhem, Ricardo Pio Monti, Diederik P. Kingma, Aapo HyvärinenNeurIPS 2020 · 被引用 141 次
- Few-shot Domain Adaptation by Causal Mechanism TransferTakeshi Teshima, Issei Sato, Masashi SugiyamaICML 2020 · 被引用 101 次
相关 Paper
- Learning Causal Semantic Representation for Out-of-Distribution PredictionChang Liu, Xinwei Sun, Jindong Wang, Haoyue Tang 等NeurIPS 2021 · 被引用 136 次
- Learning Optimal Features via Partial InvarianceMoulik Choraria, Ibtihal Ferwana, Ankur Mani, Lav R. VarshneyAAAI 2023 · 被引用 3 次
- Counterfactual Invariance to Spurious Correlations in Text ClassificationVictor Veitch, Alexander D'Amour, Steve Yadlowsky, Jacob EisensteinNeurIPS 2021 · 被引用 108 次
- Invariant and Transportable Representations for Anti-Causal Domain ShiftsYibo Jiang, Victor VeitchNeurIPS 2022 · 被引用 50 次
- Out-of-distribution Generalization with Causal Invariant TransformationsRuoyu Wang, Mingyang Yi, Zhitang Chen, Shengyu ZhuCVPR 2022 · 被引用 40 次
