A Theory of Independent Mechanisms for Extrapolation in Generative Models
Michel Besserve, Rémy Sun, Dominik Janzing, Bernhard Schölkopf
摘要
Generative models can be trained to emulate complex empirical data, but are they useful to make predictions in the context of previously unobserved environments? An intuitive idea to promote such extrapolation capabilities is to have the architecture of such model reflect a causal graph of the true data generating process, such that one can intervene on each node independently of the others. However, the nodes of this graph are usually unobserved, leading to overparameterization and lack of identifiability of the causal structure. We develop a theoretical framework to address this challenging situation by defining a weaker form of identifiability, based on the principle of independence of mechanisms. We demonstrate on toy examples that classical stochastic gradient descent can hinder the model's extrapolation capabilities, suggesting independence of mechanisms should be enforced explicitly during training. Experiments on deep generative models trained on real world data support these insights and illustrate how the extrapolation capabilities of such models can be leveraged.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper10
- Visual Representation Learning Does Not Generalize Strongly Within the Same DomainLukas Schott, Julius von Kügelgen, Frederik Träuble, Peter Vincent Gehler 等ICLR 2022 · 被引用 79 次
- Additive Decoders for Latent Variables Identification and Cartesian-Product ExtrapolationSébastien Lachapelle, Divyat Mahajan, Ioannis Mitliagkas, Simon Lacoste-JulienNeurIPS 2023 · 被引用 61 次
- Provably Learning Object-Centric RepresentationsJack Brady, Roland S. Zimmermann, Yash Sharma, Bernhard Schölkopf 等ICML 2023 · 被引用 55 次
- Dynamic Inference with Neural InterpretersNasim Rahaman, Muhammad Waleed Gondal, Shruti Joshi, Peter V. Gehler 等NeurIPS 2021 · 被引用 35 次
- Demystifying Inductive Biases for (Beta-)VAE Based ArchitecturesDominik Zietlow, Michal Rolínek, Georg MartiusICML 2021 · 被引用 24 次
它引用的顶会 Paper4
- Recurrent Independent MechanismsAnirudh Goyal, Alex Lamb, Jordan Hoffmann, Shagun Sodhani 等ICLR 2021 · 被引用 357 次
- Gradient-Based Neural DAG LearningSébastien Lachapelle, Philippe Brouillard, Tristan Deleu, Simon Lacoste-JulienICLR 2020 · 被引用 337 次
- Causal Discovery with Reinforcement LearningShengyu Zhu, Ignavier Ng, Zhitang ChenICLR 2020 · 被引用 285 次
- Counterfactuals uncover the modular structure of deep generative modelsMichel Besserve, Arash Mehrjou, Rémy Sun, Bernhard SchölkopfICLR 2020 · 被引用 109 次
相关 Paper
- Independent mechanism analysis, a new concept?Luigi Gresele, Julius von Kügelgen, Vincent Stimper, Bernhard Schölkopf 等NeurIPS 2021 · 被引用 133 次
- Identifiable Generative models for Missing Not at Random Data ImputationChao Ma, Cheng ZhangNeurIPS 2021 · 被引用 56 次
- Identifiable Exchangeable Mechanisms for Causal Structure and Representation LearningPatrik Reizinger, Siyuan Guo, Ferenc Huszár, Bernhard Schölkopf 等ICLR 2025
- Nonparametric Identifiability of Causal Representations from Unknown InterventionsJulius von Kügelgen, Michel Besserve, Wendong Liang, Luigi Gresele 等NeurIPS 2023 · 被引用 127 次
- Intervention Generalization: A View from Factor Graph ModelsGecia Bravo Hermsdorff, David S. Watson, Jialin Yu, Jakob Zeitler 等NeurIPS 2023 · 被引用 7 次
