Unbiased learning of deep generative models with structured discrete representations
Henry C. Bendekgey, Gabe Hope, Erik B. Sudderth
摘要
By composing graphical models with deep learning architectures, we learn generative models with the strengths of both frameworks. The structured variational autoencoder (SVAE) inherits structure and interpretability from graphical models, and flexible likelihoods for high-dimensional data from deep learning, but poses substantial optimization challenges. We propose novel algorithms for learning SVAEs, and are the first to demonstrate the SVAE's ability to handle multimodal uncertainty when data is missing by incorporating discrete latent variables. Our memory-efficient implicit differentiation scheme makes the SVAE tractable to learn via gradient descent, while demonstrating robustness to incomplete optimization. To more rapidly learn accurate graphical model parameters, we derive a method for computing natural gradients without manual derivations, which avoids biases found in prior work. These optimization innovations enable the first comparisons of the SVAE to state-of-the-art time series models, where the SVAE performs competitively while learning interpretable and structured discrete data representations.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper7
- Are Transformers Effective for Time Series Forecasting?Ailing Zeng, Muxi Chen, Lei Zhang, Qiang XuAAAI 2023 · 被引用 3,619 次
- NVAE: A Deep Hierarchical Variational AutoencoderArash Vahdat, Jan KautzNeurIPS 2020 · 被引用 1,141 次
- Simple and Effective VAE Training with Calibrated DecodersOleh Rybkin, Kostas Daniilidis, Sergey LevineICML 2021 · 被引用 119 次
- Very Deep VAEs Generalize Autoregressive Models and Can Outperform Them on ImagesRewon ChildICLR 2021 · 被引用 45 次
- Revisiting Structured Variational AutoencodersYixiu Zhao, Scott W. LindermanICML 2023 · 被引用 15 次
相关 Paper
- Undirected Graphical Models as Approximate PosteriorsArash Vahdat, Evgeny Andriyash, William G. MacreadyICML 2020 · 被引用 15 次
- Structured Flow Autoencoders: Learning Structured Probabilistic Representations with Flow MatchingYidan Xu, Yixin Wang, XuanLong NguyenICLR 2026
- Factorized Inference in Deep Markov Models for Incomplete Multimodal Time SeriesZhi-Xuan Tan, Harold Soh, Desmond C. OngAAAI 2020 · 被引用 32 次
- Structure by Architecture: Structured Representations without RegularizationFelix Leeb, Giulia Lanzillotta, Yashas Annadani, Michel Besserve 等ICLR 2023 · 被引用 1 次
- Analytical Probability Distributions and Exact Expectation-Maximization for Deep Generative NetworksRandall Balestriero, Sébastien Paris, Richard G. BaraniukNeurIPS 2020 · 被引用 5 次
