Generalized Doubly Reparameterized Gradient Estimators
Matthias Bauer, Andriy Mnih
Abstract
Efficient low-variance gradient estimation enabled by the reparameterization trick (RT) has been essential to the success of variational autoencoders. Doubly-reparameterized gradients (DReGs) improve on the RT for multi-sample variational bounds by applying reparameterization a second time for an additional reduction in variance. Here, we develop two generalizations of the DReGs estimator and show that they can be used to train conditional and hierarchical VAEs on image modelling tasks more effectively. First, we extend the estimator to hierarchical models with several stochastic layers by showing how to treat additional score function terms due to the hierarchical variational posterior. We then generalize DReGs to score functions of arbitrary distributions instead of just those of the sampling distribution, which makes the estimator applicable to the parameters of the prior in addition to those of the posterior.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext b6f51bed-c836-42ba-b952-3fd63e6a1737Cited by top-tier papers8
- Nested Variational InferenceHeiko Zimmermann, Hao Wu, Babak Esmaeili, Jan-Willem van de MeentNeurIPS 2021 · 26 citations
- Flow Annealed Importance Sampling BootstrapLaurence Illing Midgley, Vincent Stimper, Gregor N. C. Simm, Bernhard Schölkopf et al.ICLR 2023 · 14 citations
- Path-Gradient Estimators for Continuous Normalizing FlowsLorenz Vaitl, Kim Andrea Nicoli, Shinichi Nakajima, Pan KesselICML 2022 · 14 citations
- Multi-Sample Training for Neural Image CompressionTongda Xu, Yan Wang, Dailan He, Chenjian Gao et al.NeurIPS 2022 · 7 citations
- Fast and unified path gradient estimators for normalizing flowsLorenz Vaitl, Ludwig Winkler, Lorenz Richter, Pan KesselICLR 2024 · 6 citations
Builds on2
Related papers
- VarGrad: A Low-Variance Gradient Estimator for Variational InferenceLorenz Richter, Ayman Boustati, Nikolas Nüsken, Francisco J. R. Ruiz et al.NeurIPS 2020 · 90 citations
- Undirected Graphical Models as Approximate PosteriorsArash Vahdat, Evgeny Andriyash, William G. MacreadyICML 2020 · 15 citations
- Variational Learning of Fractional PosteriorsKian Ming A. Chai, Edwin V. BonillaICML 2025
- Posterior Matching for Arbitrary ConditioningRyan R. Strauss, Junier B. OlivaNeurIPS 2022 · 7 citations
- Rao-Blackwellised Reparameterisation GradientsKevin H. Lam, Thang Bui, George Deligiannidis, Yee Whye TehNeurIPS 2025 · 1 citation
