Stochastic Gradient Variational Inference with Price's Gradient Estimator from Bures-Wasserstein to Parameter Space
Kyurae Kim, Qiang Fu, Yian Ma, Jacob Gardner, Trevor Campbell
Abstract
For approximating a target distribution given only its unnormalized log-density, stochastic gradient-based variational inference (VI) algorithms are a popular approach. For example, Wasserstein VI (WVI) and black-box VI (BBVI) perform gradient descent in measure space (Bures-Wasserstein space) and parameter space, respectively. Previously, for the Gaussian variational family, convergence guarantees for WVI have shown superiority over existing results for black-box VI with the reparametrization gradient, suggesting the measure space approach might provide some unique benefits. In this work, however, we close this gap by obtaining identical state-of-the-art iteration complexity guarantees for both. In particular, we identify that WVI's superiority stems from the specific gradient estimator it uses, which BBVI can also leverage with minor modifications. The estimator in question is usually associated with Price's theorem and utilizes second-order information (Hessians) of the target log-density. We will refer to this as Price's gradient. On the flip side, WVI can be made more widely applicable by using the reparametrization gradient, which requires only gradients of the log-density. We empirically demonstrate that the use of Price's gradient is the major source of performance improvement.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 3d9604c3-1ea9-4b25-a1fb-5734ccbfc765Builds on17
- Variational inference via Wasserstein gradient flowsMarc Lambert, Sinho Chewi, Francis R. Bach, Silvère Bonnabel et al.NeurIPS 2022 · 123 citations
- The Wasserstein Proximal Gradient AlgorithmAdil Salim, Anna Korba, Giulia LuiseNeurIPS 2020 · 74 citations
- Averaging on the Bures-Wasserstein manifold: dimension-free convergence of gradient descentJason M. Altschuler, Sinho Chewi, Patrik Gerber, Austin J. StrommeNeurIPS 2021 · 60 citations
- Advances in Black-Box VI: Normalizing Flows, Importance Weighting, and OptimizationAbhinav Agrawal, Daniel Sheldon, Justin DomkeNeurIPS 2020 · 49 citations
- Forward-Backward Gaussian Variational Inference via JKO in the Bures-Wasserstein SpaceMichael Ziyang Diao, Krishna Balasubramanian, Sinho Chewi, Adil SalimICML 2023 · 47 citations
Related papers
- On the Convergence of Black-Box Variational InferenceKyurae Kim, Jisu Oh, Kaiwen Wu, Yi-An Ma et al.NeurIPS 2023 · 27 citations
- Provable convergence guarantees for black-box variational inferenceJustin Domke, Robert M. Gower, Guillaume GarrigosNeurIPS 2023 · 35 citations
- Variational Inference with Gaussian Score MatchingChirag Modi, Robert M. Gower, Charles Margossian, Yuling Yao et al.NeurIPS 2023 · 24 citations
- Batch and match: black-box variational inference with a score-based divergenceDiana Cai, Chirag Modi, Loucas Pillaud-Vivien, Charles Margossian et al.ICML 2024 · 18 citations
- Nearly Dimension-Independent Convergence of Mean-Field Black-Box Variational InferenceKyurae Kim, Yian Ma, Trevor Campbell, Jacob R. GardnerNeurIPS 2025 · 1 citation
