VarGrad: A Low-Variance Gradient Estimator for Variational Inference
Lorenz Richter, Ayman Boustati, Nikolas Nüsken, Francisco J. R. Ruiz, Ömer Deniz Akyildiz
摘要
We analyse the properties of an unbiased gradient estimator of the evidence lower bound (ELBO) for variational inference, based on the score function method with leave-one-out control variates. We show that this gradient estimator can be obtained using a new loss, defined as the variance of the log-ratio between the exact posterior and the variational approximation, which we call the log-variance loss. Under certain conditions, the gradient of the log-variance loss equals the gradient of the (negative) ELBO. We show theoretically that this gradient estimator, which we call VarGrad due to its connection to the log-variance loss, exhibits lower variance than the score function method in certain settings, and that the leave-one-out control variate coefficients are close to the optimal ones. We empirically demonstrate that VarGrad offers a favourable variance versus computation trade-off compared to other state-of-the-art estimators on a discrete variational autoencoder (VAE).
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper44
- Generalized Preference Optimization: A Unified Approach to Offline AlignmentYunhao Tang, Zhaohan Daniel Guo, Zeyu Zheng, Daniele Calandriello 等ICML 2024 · 被引用 159 次
- REBEL: Reinforcement Learning via Regressing Relative RewardsZhaolin Gao, Jonathan D. Chang, Wenhao Zhan, Owen Oertell 等NeurIPS 2024 · 被引用 82 次
- Amortizing intractable inference in diffusion models for vision, language, and controlSiddarth Venkatraman, Moksh Jain, Luca Scimeca, Minsu Kim 等NeurIPS 2024 · 被引用 79 次
- DEFT: Efficient Fine-tuning of Diffusion Models by Learning the Generalised -transformAlexander Denker, Francisco Vargas, Shreyas Padhy, Kieran Didi 等NeurIPS 2024 · 被引用 52 次
- Improved off-policy training of diffusion samplersMarcin Sendera, Minsu Kim, Sarthak Mittal, Pablo Lemos 等NeurIPS 2024 · 被引用 52 次
它引用的顶会 Paper3
- Markovian Score Climbing: Variational Inference with KL(p||q)Christian A. Naesseth, Fredrik Lindsten, David M. BleiNeurIPS 2020 · 被引用 67 次
- Estimating Gradients for Discrete Random Variables by Sampling without ReplacementWouter Kool, Herke van Hoof, Max WellingICLR 2020 · 被引用 59 次
- DisARM: An Antithetic Gradient Estimator for Binary Latent VariablesZhe Dong, Andriy Mnih, George TuckerNeurIPS 2020 · 被引用 43 次
相关 Paper
- Gradient Estimation with Discrete Stein OperatorsJiaxin Shi, Yuhao Zhou, Jessica Hwang, Michalis K. Titsias 等NeurIPS 2022 · 被引用 27 次
- SUMO: Unbiased Estimation of Log Marginal Probability for Latent Variable ModelsYucen Luo, Alex Beatson, Mohammad Norouzi, Jun Zhu 等ICLR 2020 · 被引用 29 次
- Quantized Variational InferenceAmir DibNeurIPS 2020 · 被引用 1 次
- Generalized Doubly Reparameterized Gradient EstimatorsMatthias Bauer, Andriy MnihICML 2021 · 被引用 15 次
- Variational (Gradient) Estimate of the Score Function in Energy-based Latent Variable ModelsFan Bao, Kun Xu, Chongxuan Li, Lanqing Hong 等ICML 2021 · 被引用 10 次
