Revisiting Structured Variational Autoencoders
Yixiu Zhao, Scott W. Linderman
摘要
Structured variational autoencoders (SVAEs) (Johnson et al., 2016) combine probabilistic graphical model priors on latent variables, deep neural networks to link latent variables to observed data, and structure-exploiting algorithms for approximate posterior inference. These models are particularly appealing for sequential data, where the prior can capture temporal dependencies. However, despite their conceptual elegance, SVAEs have proven difficult to implement, and more general approaches have been favored in practice. Here, we revisit SVAEs using modern machine learning tools and demonstrate their advantages over more general alternatives in terms of both accuracy and efficiency. First, we develop a modern implementation for hardware acceleration, parallelization, and automatic differentiation of the message passing algorithms at the core of the SVAE. Second, we show that by exploiting structure in the prior, the SVAE learns more accurate models and posterior distributions, which translate into improved performance on prediction tasks. Third, we show how the SVAE can naturally handle missing data, and we leverage this ability to develop a novel, self-supervised training approach. Altogether, these results show that the time is ripe to revisit structured variational autoencoders.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper6
- eXponential FAmily Dynamical Systems (XFADS): Large-scale nonlinear Gaussian state-space modelingMatthew Dowling, Yuan Zhao, Il Memming ParkNeurIPS 2024 · 被引用 17 次
- Modeling state-dependent communication between brain regions with switching nonlinear dynamical systemsOrren Karniol-Tambour, David M. Zoltowski, E. Mika Diamanti, Lucas Pinto 等ICLR 2024 · 被引用 14 次
- Unbiased learning of deep generative models with structured discrete representationsHenry C. Bendekgey, Gabe Hope, Erik B. SudderthNeurIPS 2023 · 被引用 2 次
- Maximum-Likelihood Learning of Latent Dynamics Without ReconstructionSamo Hromadka, Kai Biegun, Lior Fox, James Heald 等ICML 2026
- Efficient Learning of Deep State Space Models via Importance SmoothingJohn-Joseph Brady, Nikolas Nüsken, Yunpeng LiICML 2026
它引用的顶会 Paper10
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah 等NeurIPS 2020 · 被引用 64,255 次
- Efficiently Modeling Long Sequences with Structured State SpacesAlbert Gu, Karan Goel, Christopher RéICLR 2022 · 被引用 3,482 次
- Diagonal State Spaces are as Effective as Structured State SpacesAnkit Gupta, Albert Gu, Jonathan BerantNeurIPS 2022 · 被引用 546 次
- Efficient and Modular Implicit DifferentiationMathieu Blondel, Quentin Berthet, Marco Cuturi, Roy Frostig 等NeurIPS 2022 · 被引用 386 次
- Modeling Irregular Time Series with Continuous Recurrent UnitsMona Schirmer, Mazin Eltayeb, Stefan Lessmann, Maja RudolphICML 2022 · 被引用 135 次
相关 Paper
- From Variational to Deterministic AutoencodersPartha Ghosh, Mehdi S. M. Sajjadi, Antonio Vergari, Michael J. Black 等ICLR 2020 · 被引用 298 次
- Sparse Autoencoders, Again?Yin Lu, Xuening Zhu, Tong He, David WipfICML 2025
- Structure by Architecture: Structured Representations without RegularizationFelix Leeb, Giulia Lanzillotta, Yashas Annadani, Michel Besserve 等ICLR 2023 · 被引用 1 次
- Undirected Graphical Models as Approximate PosteriorsArash Vahdat, Evgeny Andriyash, William G. MacreadyICML 2020 · 被引用 15 次
- Latent variable model for high-dimensional point process with structured missingnessMaksim Sinelnikov, Manuel Haussmann, Harri LähdesmäkiICML 2024 · 被引用 1 次
