The Usual Suspects? Reassessing Blame for VAE Posterior Collapse
Bin Dai, Ziyu Wang, David P. Wipf
Abstract
In narrow asymptotic settings Gaussian VAE models of continuous data have been shown to possess global optima aligned with ground-truth distributions. Even so, it is well known that poor solutions whereby the latent posterior collapses to an uninformative prior are sometimes obtained in practice. However, contrary to conventional wisdom that largely assigns blame for this phenomena on the undue influence of KL-divergence regularization, we will argue that posterior collapse is, at least in part, a direct consequence of bad local minima inherent to the loss surface of deep autoencoder networks. In particular, we prove that even small nonlinear perturbations of affine VAE decoder models can produce such minima, and in deeper models, analogous minima can force the VAE to behave like an aggressive truncation operator, provably discarding information along all latent dimensions in certain circumstances. Regardless, the underlying message here is not meant to undercut valuable existing explanations of posterior collapse, but rather, to refine the discussion and elucidate alternative risk factors that may have been previously underappreciated.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 289264d5-1866-4259-80e5-1a8573f433efCited by top-tier papers22
- Posterior Collapse and Latent Variable Non-identifiabilityYixin Wang, David M. Blei, John P. CunninghamNeurIPS 2021 · 97 citations
- Identifiability of deep generative models without auxiliary informationBohdan Kivva, Goutham Rajendran, Pradeep Ravikumar, Bryon AragamNeurIPS 2022 · 87 citations
- A Critical Look at the Consistency of Causal Estimation with Deep Latent Variable ModelsSeveri Rissanen, Pekka MarttinenNeurIPS 2021 · 38 citations
- From Causal to Concept-Based Representation LearningGoutham Rajendran, Simon Buchholz, Bryon Aragam, Bernhard Schölkopf et al.NeurIPS 2024 · 37 citations
- Embrace the Gap: VAEs Perform Independent Mechanism AnalysisPatrik Reizinger, Luigi Gresele, Jack Brady, Julius von Kügelgen et al.NeurIPS 2022 · 34 citations
Related papers
- Posterior Collapse of a Linear Latent Variable ModelZihao Wang, Liu ZiyinNeurIPS 2022 · 29 citations
- Marginalization is not Marginal: No Bad VAE Local Minima when Learning Optimal Sparse RepresentationsDavid WipfICML 2023 · 5 citations
- Beyond Vanilla Variational Autoencoders: Detecting Posterior Collapse in Conditional and Hierarchical Variational AutoencodersHien Dang, Tho Tran Huu, Tan Minh Nguyen, Nhat HoICLR 2024 · 8 citations
- Spectral Smoothing Unveils Phase Transitions in Hierarchical Variational AutoencodersAdeel Pervez, Efstratios GavvesICML 2021 · 4 citations
- On the Value of Infinite Gradients in Variational Autoencoder ModelsBin Dai, Wenliang Li, David P. WipfNeurIPS 2021 · 15 citations
