VarGrad: A Low-Variance Gradient Estimator for Variational Inference
Lorenz Richter, Ayman Boustati, Nikolas Nüsken, Francisco J. R. Ruiz, Ömer Deniz Akyildiz
Abstract
We analyse the properties of an unbiased gradient estimator of the evidence lower bound (ELBO) for variational inference, based on the score function method with leave-one-out control variates. We show that this gradient estimator can be obtained using a new loss, defined as the variance of the log-ratio between the exact posterior and the variational approximation, which we call the log-variance loss. Under certain conditions, the gradient of the log-variance loss equals the gradient of the (negative) ELBO. We show theoretically that this gradient estimator, which we call VarGrad due to its connection to the log-variance loss, exhibits lower variance than the score function method in certain settings, and that the leave-one-out control variate coefficients are close to the optimal ones. We empirically demonstrate that VarGrad offers a favourable variance versus computation trade-off compared to other state-of-the-art estimators on a discrete variational autoencoder (VAE).
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 89d78fa3-128e-416d-9fd8-aa6da504f970Cited by top-tier papers44
- Generalized Preference Optimization: A Unified Approach to Offline AlignmentYunhao Tang, Zhaohan Daniel Guo, Zeyu Zheng, Daniele Calandriello et al.ICML 2024 · 159 citations
- REBEL: Reinforcement Learning via Regressing Relative RewardsZhaolin Gao, Jonathan D. Chang, Wenhao Zhan, Owen Oertell et al.NeurIPS 2024 · 82 citations
- Amortizing intractable inference in diffusion models for vision, language, and controlSiddarth Venkatraman, Moksh Jain, Luca Scimeca, Minsu Kim et al.NeurIPS 2024 · 79 citations
- DEFT: Efficient Fine-tuning of Diffusion Models by Learning the Generalised -transformAlexander Denker, Francisco Vargas, Shreyas Padhy, Kieran Didi et al.NeurIPS 2024 · 52 citations
- Improved off-policy training of diffusion samplersMarcin Sendera, Minsu Kim, Sarthak Mittal, Pablo Lemos et al.NeurIPS 2024 · 52 citations
Builds on3
- Markovian Score Climbing: Variational Inference with KL(p||q)Christian A. Naesseth, Fredrik Lindsten, David M. BleiNeurIPS 2020 · 67 citations
- Estimating Gradients for Discrete Random Variables by Sampling without ReplacementWouter Kool, Herke van Hoof, Max WellingICLR 2020 · 59 citations
- DisARM: An Antithetic Gradient Estimator for Binary Latent VariablesZhe Dong, Andriy Mnih, George TuckerNeurIPS 2020 · 43 citations
Related papers
- Gradient Estimation with Discrete Stein OperatorsJiaxin Shi, Yuhao Zhou, Jessica Hwang, Michalis K. Titsias et al.NeurIPS 2022 · 27 citations
- SUMO: Unbiased Estimation of Log Marginal Probability for Latent Variable ModelsYucen Luo, Alex Beatson, Mohammad Norouzi, Jun Zhu et al.ICLR 2020 · 29 citations
- Quantized Variational InferenceAmir DibNeurIPS 2020 · 1 citation
- Generalized Doubly Reparameterized Gradient EstimatorsMatthias Bauer, Andriy MnihICML 2021 · 15 citations
- Variational (Gradient) Estimate of the Score Function in Energy-based Latent Variable ModelsFan Bao, Kun Xu, Chongxuan Li, Lanqing Hong et al.ICML 2021 · 10 citations
