Optimal Variance Control of the Score-Function Gradient Estimator for Importance-Weighted Bounds
Valentin Liévin, Andrea Dittadi, Anders Christensen, Ole Winther
Abstract
This paper introduces novel results for the score function gradient estimator of the importance weighted variational bound (IWAE). We prove that in the limit of large K (number of importance samples) one can choose the control variate such that the Signal-to-Noise ratio (SNR) of the estimator grows as √ K. This is in contrast to the standard pathwise gradient estimator where the SNR decreases as 1/ √ K. Based on our theoretical findings we develop a novel control variate that extends on VIMCO. Empirically, for the training of both continuous and discrete generative models, the proposed method yields superior variance reduction, resulting in an SNR for IWAE that increases with K without relying on the reparameterization trick. The novel estimator is competitive with state-of-the-art reparameterization-free gradient estimators such as Reweighted Wake-Sleep (RWS) and the thermodynamic variational objective (TVO) when training generative models. Recently, variational objectives tighter than the traditional evidence lower bound (ELBO) have been proposed [21, 22] . In importance weighted autoencoders (IWAE) [22] the tighter bound comes with the price of a K-fold increase in the required number of samples from the inference network. Despite yielding a tighter bound, using more samples can be detrimental to the learning of the inference model [23] . In fact, the Signal-to-Noise ratio (the ratio of the expected gradient to its standard deviation) of the pathwise estimator has been shown to decrease at a rate O(K -1/2 ) [23] . Although this can be improved to O(K 1/2 ) by exploiting properties of the gradient to cancel high-variance 34th Conference on Neural Information Processing Systems (NeurIPS 2020),
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers3
- Planning from Pixels in Atari with Learned Symbolic RepresentationsAndrea Dittadi, Frederik K. Drachmann, Thomas BolanderAAAI 2021 · 14 citations
- Variational Open-Domain Question AnsweringValentin Liévin, Andreas Geert Motzfeldt, Ida Riis Jensen, Ole WintherICML 2023 · 11 citations
- Understanding the difficulties of posterior predictive estimationAbhinav Agrawal, Justin DomkeICML 2025
Related papers
- VarGrad: A Low-Variance Gradient Estimator for Variational InferenceLorenz Richter, Ayman Boustati, Nikolas Nüsken, Francisco J. R. Ruiz et al.NeurIPS 2020 · 90 citations
- All in the Exponential Family: Bregman Duality in Thermodynamic Variational InferenceRob Brekelmans, Vaden Masrani, Frank Wood, Greg Ver Steeg et al.ICML 2020 · 18 citations
- Multi-Sample Training for Neural Image CompressionTongda Xu, Yan Wang, Dailan He, Chenjian Gao et al.NeurIPS 2022 · 7 citations
- SUMO: Unbiased Estimation of Log Marginal Probability for Latent Variable ModelsYucen Luo, Alex Beatson, Mohammad Norouzi, Jun Zhu et al.ICLR 2020 · 29 citations
- Improving Mutual Information Estimation with Annealed and Energy-Based BoundsRob Brekelmans, Sicong Huang, Marzyeh Ghassemi, Greg Ver Steeg et al.ICLR 2022 · 16 citations
