Posterior Collapse and Latent Variable Non-identifiability
Yixin Wang, David M. Blei, John P. Cunningham
摘要
Variational autoencoders model high-dimensional data by positing low-dimensional latent variables that are mapped through a flexible distribution parametrized by a neural network. Unfortunately, variational autoencoders often suffer from posterior collapse: the posterior of the latent variables is equal to its prior, rendering the variational autoencoder useless as a means to produce meaningful representations. Existing approaches to posterior collapse often attribute it to the use of neural networks or optimization issues due to variational approximation. In this paper, we consider posterior collapse as a problem of latent variable non-identifiability. We prove that the posterior collapses if and only if the latent variables are non-identifiable in the generative model. This fact implies that posterior collapse is not a phenomenon specific to the use of flexible distributions or approximate inference. Rather, it can occur in classical probabilistic models even with exact inference, which we also demonstrate. Based on these results, we propose a class of latent-identifiable variational autoencoders, deep generative models which enforce identifiability without sacrificing flexibility. This model class resolves the problem of latent variable non-identifiability by leveraging bijective Brenier maps and parameterizing them with input convex neural networks, without special variational inference objectives or optimization tricks. Across synthetic and real datasets, latent-identifiable variational autoencoders outperform existing methods in mitigating posterior collapse and providing meaningful representations of the data.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper31
- Learning Linear Causal Representations from Interventions under General Nonlinear MixingSimon Buchholz, Goutham Rajendran, Elan Rosenfeld, Bryon Aragam 等NeurIPS 2023 · 被引用 113 次
- Identifiability of deep generative models without auxiliary informationBohdan Kivva, Goutham Rajendran, Pradeep Ravikumar, Bryon AragamNeurIPS 2022 · 被引用 87 次
- GFlowNet-EM for Learning Compositional Latent Variable ModelsEdward J. Hu, Nikolay Malkin, Moksh Jain, Katie E. Everett 等ICML 2023 · 被引用 48 次
- Learning Nonparametric Latent Causal Graphs with Unknown InterventionsYibo Jiang, Bryon AragamNeurIPS 2023 · 被引用 39 次
- From Causal to Concept-Based Representation LearningGoutham Rajendran, Simon Buchholz, Bryon Aragam, Bernhard Schölkopf 等NeurIPS 2024 · 被引用 37 次
它引用的顶会 Paper4
- Optimal transport mapping via input convex neural networksAshok Vardhan Makkuva, Amirhossein Taghvaei, Sewoong Oh, Jason D. LeeICML 2020 · 被引用 254 次
- ICE-BeeM: Identifiable Conditional Energy-Based Deep Models Based on Nonlinear ICAIlyes Khemakhem, Ricardo Pio Monti, Diederik P. Kingma, Aapo HyvärinenNeurIPS 2020 · 被引用 141 次
- The Usual Suspects? Reassessing Blame for VAE Posterior CollapseBin Dai, Ziyu Wang, David P. WipfICML 2020 · 被引用 89 次
- On Implicit Regularization in β-VAEsAbhishek Kumar, Ben PooleICML 2020 · 被引用 59 次
相关 Paper
- Posterior Collapse of a Linear Latent Variable ModelZihao Wang, Liu ZiyinNeurIPS 2022 · 被引用 29 次
- Controlling Posterior Collapse by an Inverse Lipschitz Constraint on the Decoder NetworkYuri Kinoshita, Kenta Oono, Kenji Fukumizu, Yuichi Yoshida 等ICML 2023 · 被引用 6 次
- Beyond Vanilla Variational Autoencoders: Detecting Posterior Collapse in Conditional and Hierarchical Variational AutoencodersHien Dang, Tho Tran Huu, Tan Minh Nguyen, Nhat HoICLR 2024 · 被引用 8 次
- Effective Estimation of Deep Generative Language ModelsTom Pelsmaeker, Wilker AzizACL 2020 · 被引用 5 次
- Spectral Smoothing Unveils Phase Transitions in Hierarchical Variational AutoencodersAdeel Pervez, Efstratios GavvesICML 2021 · 被引用 4 次
