Variational autoencoders in the presence of low-dimensional data: landscape and implicit bias
Frederic Koehler, Viraj Mehta, Chenghui Zhou, Andrej Risteski
Abstract
Variational Autoencoders (VAEs) are one of the most commonly used generative models, particularly for image data. A prominent difficulty in training VAEs is data that is supported on a lower dimensional manifold. Recent work by Dai and Wipf (2020) proposes a two-stage training algorithm for VAEs, based on a conjecture that in standard VAE training the generator will converge to a solution with 0 variance which is correctly supported on the ground truth manifold. They gave partial support for this conjecture by showing that some optima of the VAE loss do satisfy this property, but did not analyze the training dynamics. In this paper, we show that for linear encoders/decoders, the conjecture is true-that is the VAE training does recover a generator with support equal to the ground truth manifold-and does so due to an implicit bias of gradient descent rather than merely the VAE loss itself. In the nonlinear case, we show that VAE training frequently learns a higher-dimensional manifold which is a superset of the ground truth manifold.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext b4fe0aa9-9a02-4011-9e38-9bdf6ea1b6acCited by top-tier papers3
- Marginalization is not Marginal: No Bad VAE Local Minima when Learning Optimal Sparse RepresentationsDavid WipfICML 2023 · 5 citations
- A Statistical Analysis of Wasserstein Autoencoders for Intrinsically Low-dimensional DataSaptarshi Chakraborty, Peter L. BartlettICLR 2024 · 3 citations
- Be a Goldfish: Forgetting Bad Conditioning in Sparse Linear Regression via Variational AutoencodersKuheli Pratihar, Debdeep MukhopadhyayICML 2025
Builds on5
- NVAE: A Deep Hierarchical Variational AutoencoderArash Vahdat, Jan KautzNeurIPS 2020 · 1,141 citations
- Towards Resolving the Implicit Bias of Gradient Descent for Matrix Factorization: Greedy Low-Rank LearningZhiyuan Li, Yuping Luo, Kaifeng LyuICLR 2021 · 155 citations
- The Usual Suspects? Reassessing Blame for VAE Posterior CollapseBin Dai, Ziyu Wang, David P. WipfICML 2020 · 89 citations
- On Implicit Regularization in β-VAEsAbhishek Kumar, Ben PooleICML 2020 · 59 citations
- Spectral Smoothing Unveils Phase Transitions in Hierarchical Variational AutoencodersAdeel Pervez, Efstratios GavvesICML 2021 · 4 citations
Related papers
- Learning Manifold Dimensions with Conditional Variational AutoencodersYijia Zheng, Tong He, Yixuan Qiu, David P. WipfNeurIPS 2022 · 34 citations
- On the Value of Infinite Gradients in Variational Autoencoder ModelsBin Dai, Wenliang Li, David P. WipfNeurIPS 2021 · 15 citations
- A solvable model of learning generative diffusion: theory and insightsHugo Cui, Cengiz Pehlevan, Yue M. LuNeurIPS 2025 · 11 citations
- The Effects of Invertibility on the Representational Complexity of Encoders in Variational AutoencodersDivyansh Pareek, Andrej RisteskiICLR 2022
- Embrace the Gap: VAEs Perform Independent Mechanism AnalysisPatrik Reizinger, Luigi Gresele, Jack Brady, Julius von Kügelgen et al.NeurIPS 2022 · 34 citations
