Additive Decoders for Latent Variables Identification and Cartesian-Product Extrapolation
Sébastien Lachapelle, Divyat Mahajan, Ioannis Mitliagkas, Simon Lacoste-Julien
Abstract
We tackle the problems of latent variables identification and ``out-of-support'' image generation in representation learning. We show that both are possible for a class of decoders that we call additive, which are reminiscent of decoders used for object-centric representation learning (OCRL) and well suited for images that can be decomposed as a sum of object-specific images. We provide conditions under which exactly solving the reconstruction problem using an additive decoder is guaranteed to identify the blocks of latent variables up to permutation and block-wise invertible transformations. This guarantee relies only on very weak assumptions about the distribution of the latent factors, which might present statistical dependencies and have an almost arbitrarily shaped support. Our result provides a new setting where nonlinear independent component analysis (ICA) is possible and adds to our theoretical understanding of OCRL methods. We also show theoretically that additive decoders can generate novel images by recombining observed factors of variations in novel ways, an ability we refer to as Cartesian-product extrapolation. We show empirically that additivity is crucial for both identifiability and extrapolation on simulated data.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 31851571-4aa4-4124-94f4-c7ed5693fb88Cited by top-tier papers33
- Provable Compositional Generalization for Object-Centric LearningThaddäus Wiedemer, Jack Brady, Alexander Panfilov, Attila Juhos et al.ICLR 2024 · 40 citations
- Marrying Causal Representation Learning with Dynamical Systems for ScienceDingling Yao, Caroline Muller, Francesco LocatelloNeurIPS 2024 · 29 citations
- Object centric architectures enable efficient causal representation learningAmin Mansouri, Jason S. Hartford, Yan Zhang, Yoshua BengioICLR 2024 · 28 citations
- CaRiNG: Learning Temporal Causal Representation under Non-Invertible Generation ProcessGuangyi Chen, Yifan Shen, Zhenhao Chen, Xiangchen Song et al.ICML 2024 · 22 citations
- Learning Discrete Concepts in Latent Hierarchical ModelsLingjing Kong, Guangyi Chen, Biwei Huang, Eric P. Xing et al.NeurIPS 2024 · 20 citations
Builds on31
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- Object-Centric Learning with Slot AttentionFrancesco Locatello, Dirk Weissenborn, Thomas Unterthiner, Aravindh Mahendran et al.NeurIPS 2020 · 1,275 citations
- Out-of-Distribution Generalization via Risk Extrapolation (REx)David Krueger, Ethan Caballero, Jörn-Henrik Jacobsen, Amy Zhang et al.ICML 2021 · 1,163 citations
- Self-Supervised Learning with Data Augmentations Provably Isolates Content from StyleJulius von Kügelgen, Yash Sharma, Luigi Gresele, Wieland Brendel et al.NeurIPS 2021 · 421 citations
- Weakly-Supervised Disentanglement Without CompromisesFrancesco Locatello, Ben Poole, Gunnar Rätsch, Bernhard Schölkopf et al.ICML 2020 · 361 citations
Related papers
- Causal Component AnalysisWendong Liang, Armin Kekic, Julius von Kügelgen, Simon Buchholz et al.NeurIPS 2023 · 65 citations
- On the Identifiability of Nonlinear ICA: Sparsity and BeyondYujia Zheng, Ignavier Ng, Kun ZhangNeurIPS 2022 · 104 citations
- Robustness of Nonlinear Representation LearningSimon Buchholz, Bernhard SchölkopfICML 2024 · 11 citations
- Identifying Representations for Intervention ExtrapolationSorawit Saengkyongam, Elan Rosenfeld, Pradeep Kumar Ravikumar, Niklas Pfister et al.ICLR 2024 · 20 citations
- Independent mechanism analysis, a new concept?Luigi Gresele, Julius von Kügelgen, Vincent Stimper, Bernhard Schölkopf et al.NeurIPS 2021 · 133 citations
