Optimal Variance Control of the Score-Function Gradient Estimator for Importance-Weighted Bounds
Valentin Liévin, Andrea Dittadi, Anders Christensen, Ole Winther
摘要
This paper introduces novel results for the score function gradient estimator of the importance weighted variational bound (IWAE). We prove that in the limit of large K (number of importance samples) one can choose the control variate such that the Signal-to-Noise ratio (SNR) of the estimator grows as √ K. This is in contrast to the standard pathwise gradient estimator where the SNR decreases as 1/ √ K. Based on our theoretical findings we develop a novel control variate that extends on VIMCO. Empirically, for the training of both continuous and discrete generative models, the proposed method yields superior variance reduction, resulting in an SNR for IWAE that increases with K without relying on the reparameterization trick. The novel estimator is competitive with state-of-the-art reparameterization-free gradient estimators such as Reweighted Wake-Sleep (RWS) and the thermodynamic variational objective (TVO) when training generative models. Recently, variational objectives tighter than the traditional evidence lower bound (ELBO) have been proposed [21, 22] . In importance weighted autoencoders (IWAE) [22] the tighter bound comes with the price of a K-fold increase in the required number of samples from the inference network. Despite yielding a tighter bound, using more samples can be detrimental to the learning of the inference model [23] . In fact, the Signal-to-Noise ratio (the ratio of the expected gradient to its standard deviation) of the pathwise estimator has been shown to decrease at a rate O(K -1/2 ) [23] . Although this can be improved to O(K 1/2 ) by exploiting properties of the gradient to cancel high-variance 34th Conference on Neural Information Processing Systems (NeurIPS 2020),
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- Planning from Pixels in Atari with Learned Symbolic RepresentationsAndrea Dittadi, Frederik K. Drachmann, Thomas BolanderAAAI 2021 · 被引用 14 次
- Variational Open-Domain Question AnsweringValentin Liévin, Andreas Geert Motzfeldt, Ida Riis Jensen, Ole WintherICML 2023 · 被引用 11 次
- Understanding the difficulties of posterior predictive estimationAbhinav Agrawal, Justin DomkeICML 2025
相关 Paper
- VarGrad: A Low-Variance Gradient Estimator for Variational InferenceLorenz Richter, Ayman Boustati, Nikolas Nüsken, Francisco J. R. Ruiz 等NeurIPS 2020 · 被引用 90 次
- All in the Exponential Family: Bregman Duality in Thermodynamic Variational InferenceRob Brekelmans, Vaden Masrani, Frank Wood, Greg Ver Steeg 等ICML 2020 · 被引用 18 次
- Multi-Sample Training for Neural Image CompressionTongda Xu, Yan Wang, Dailan He, Chenjian Gao 等NeurIPS 2022 · 被引用 7 次
- SUMO: Unbiased Estimation of Log Marginal Probability for Latent Variable ModelsYucen Luo, Alex Beatson, Mohammad Norouzi, Jun Zhu 等ICLR 2020 · 被引用 29 次
- Improving Mutual Information Estimation with Annealed and Energy-Based BoundsRob Brekelmans, Sicong Huang, Marzyeh Ghassemi, Greg Ver Steeg 等ICLR 2022 · 被引用 16 次
