Embrace the Gap: VAEs Perform Independent Mechanism Analysis
Patrik Reizinger, Luigi Gresele, Jack Brady, Julius von Kügelgen, Dominik Zietlow, Bernhard Schölkopf, Georg Martius, Wieland Brendel, Michel Besserve
Abstract
Variational autoencoders (VAEs) are a popular framework for modeling complex data distributions; they can be efficiently trained via variational inference by maximizing the evidence lower bound (ELBO), at the expense of a gap to the exact (log-)marginal likelihood. While VAEs are commonly used for disentangled representation learning, it is unclear why ELBO maximization would yield such representations, since unregularized maximum likelihood estimation generally cannot invert the data-generating process without additional assumptions. Yet, VAEs often succeed at this task. We seek to elucidate this apparent paradox by studying nonlinear VAEs in the limit of near-deterministic decoders. We first prove that, in this regime, the optimal encoder approximately inverts the decoder-a commonly used but unproven conjecture-which we refer to as self-consistency. Leveraging self-consistency, we show that the ELBO converges to a regularized log-likelihood. This allows VAEs to perform what has recently been termed independent mechanism analysis (IMA): it adds an inductive bias towards decoders with column-orthogonal Jacobians, which helps recovering the true latent factors. The gap between ELBO and log-likelihood is therefore welcome, since it bears unanticipated benefits for nonlinear representation learning. In experiments on synthetic and image data, we show that VAEs uncover the true latent factors when the data generating process satisfies the IMA assumption.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext dcd0c737-61fe-4d7a-836b-f2b17c2920b1Cited by top-tier papers9
- Compositional Generalization from First PrinciplesThaddäus Wiedemer, Prasanna Mayilvahanan, Matthias Bethge, Wieland BrendelNeurIPS 2023 · 78 citations
- High Fidelity Image Counterfactuals with Probabilistic Causal ModelsFabio De Sousa Ribeiro, Tian Xia, Miguel Monteiro, Nick Pawlowski et al.ICML 2023 · 68 citations
- Additive Decoders for Latent Variables Identification and Cartesian-Product ExtrapolationSébastien Lachapelle, Divyat Mahajan, Ioannis Mitliagkas, Simon Lacoste-JulienNeurIPS 2023 · 61 citations
- Provably Learning Object-Centric RepresentationsJack Brady, Roland S. Zimmermann, Yash Sharma, Bernhard Schölkopf et al.ICML 2023 · 55 citations
- Robustness of Nonlinear Representation LearningSimon Buchholz, Bernhard SchölkopfICML 2024 · 11 citations
Builds on14
- On Mutual Information Maximization for Representation LearningMichael Tschannen, Josip Djolonga, Paul K. Rubenstein, Sylvain Gelly et al.ICLR 2020 · 559 citations
- Weakly-Supervised Disentanglement Without CompromisesFrancesco Locatello, Ben Poole, Gunnar Rätsch, Bernhard Schölkopf et al.ICML 2020 · 361 citations
- From Variational to Deterministic AutoencodersPartha Ghosh, Mehdi S. M. Sajjadi, Antonio Vergari, Michael J. Black et al.ICLR 2020 · 298 citations
- Towards Nonlinear Disentanglement in Natural Data with Temporal Sparse CodingDavid A. Klindt, Lukas Schott, Yash Sharma, Ivan Ustyuzhaninov et al.ICLR 2021 · 156 citations
- On the Pitfalls of Heteroscedastic Uncertainty Estimation with Probabilistic Neural NetworksMaximilian Seitzer, Arash Tavakoli, Dimitrije Antic, Georg MartiusICLR 2022 · 122 citations
Related papers
- Why do Variational Autoencoders Really Promote Disentanglement?Pratik Bhowal, Achint Soni, Sirisha RambhatlaICML 2024 · 11 citations
- Unsupervised Disentanglement Without Compromises : How Functional Orthogonality Enforces IdentifiabilityMathieu Simon, Pascal Frossard, Christophe De VleeschouwerICML 2026
- Identifying through Flows for Recovering Latent RepresentationsShen Li, Bryan Hooi, Gim Hee LeeICLR 2020 · 15 citations
- The Autoencoding Variational AutoencoderA. Taylan Cemgil, Sumedh Ghaisas, Krishnamurthy Dvijotham, Sven Gowal et al.NeurIPS 2020 · 81 citations
- Structure by Architecture: Structured Representations without RegularizationFelix Leeb, Giulia Lanzillotta, Yashas Annadani, Michel Besserve et al.ICLR 2023 · 1 citation
