Revisiting Structured Variational Autoencoders
Yixiu Zhao, Scott W. Linderman
Abstract
Structured variational autoencoders (SVAEs) (Johnson et al., 2016) combine probabilistic graphical model priors on latent variables, deep neural networks to link latent variables to observed data, and structure-exploiting algorithms for approximate posterior inference. These models are particularly appealing for sequential data, where the prior can capture temporal dependencies. However, despite their conceptual elegance, SVAEs have proven difficult to implement, and more general approaches have been favored in practice. Here, we revisit SVAEs using modern machine learning tools and demonstrate their advantages over more general alternatives in terms of both accuracy and efficiency. First, we develop a modern implementation for hardware acceleration, parallelization, and automatic differentiation of the message passing algorithms at the core of the SVAE. Second, we show that by exploiting structure in the prior, the SVAE learns more accurate models and posterior distributions, which translate into improved performance on prediction tasks. Third, we show how the SVAE can naturally handle missing data, and we leverage this ability to develop a novel, self-supervised training approach. Altogether, these results show that the time is ripe to revisit structured variational autoencoders.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext d52bb1da-1b1a-4605-b704-3e1022aa8990Cited by top-tier papers6
- eXponential FAmily Dynamical Systems (XFADS): Large-scale nonlinear Gaussian state-space modelingMatthew Dowling, Yuan Zhao, Il Memming ParkNeurIPS 2024 · 17 citations
- Modeling state-dependent communication between brain regions with switching nonlinear dynamical systemsOrren Karniol-Tambour, David M. Zoltowski, E. Mika Diamanti, Lucas Pinto et al.ICLR 2024 · 14 citations
- Unbiased learning of deep generative models with structured discrete representationsHenry C. Bendekgey, Gabe Hope, Erik B. SudderthNeurIPS 2023 · 2 citations
- Maximum-Likelihood Learning of Latent Dynamics Without ReconstructionSamo Hromadka, Kai Biegun, Lior Fox, James Heald et al.ICML 2026
- Efficient Learning of Deep State Space Models via Importance SmoothingJohn-Joseph Brady, Nikolas Nüsken, Yunpeng LiICML 2026
Builds on10
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah et al.NeurIPS 2020 · 64,255 citations
- Efficiently Modeling Long Sequences with Structured State SpacesAlbert Gu, Karan Goel, Christopher RéICLR 2022 · 3,482 citations
- Diagonal State Spaces are as Effective as Structured State SpacesAnkit Gupta, Albert Gu, Jonathan BerantNeurIPS 2022 · 546 citations
- Efficient and Modular Implicit DifferentiationMathieu Blondel, Quentin Berthet, Marco Cuturi, Roy Frostig et al.NeurIPS 2022 · 386 citations
- Modeling Irregular Time Series with Continuous Recurrent UnitsMona Schirmer, Mazin Eltayeb, Stefan Lessmann, Maja RudolphICML 2022 · 135 citations
Related papers
- From Variational to Deterministic AutoencodersPartha Ghosh, Mehdi S. M. Sajjadi, Antonio Vergari, Michael J. Black et al.ICLR 2020 · 298 citations
- Sparse Autoencoders, Again?Yin Lu, Xuening Zhu, Tong He, David WipfICML 2025
- Structure by Architecture: Structured Representations without RegularizationFelix Leeb, Giulia Lanzillotta, Yashas Annadani, Michel Besserve et al.ICLR 2023 · 1 citation
- Undirected Graphical Models as Approximate PosteriorsArash Vahdat, Evgeny Andriyash, William G. MacreadyICML 2020 · 15 citations
- Latent variable model for high-dimensional point process with structured missingnessMaksim Sinelnikov, Manuel Haussmann, Harri LähdesmäkiICML 2024 · 1 citation
