Forward-Backward Gaussian Variational Inference via JKO in the Bures-Wasserstein Space
Michael Ziyang Diao, Krishna Balasubramanian, Sinho Chewi, Adil Salim
Abstract
Variational inference (VI) seeks to approximate a target distribution by an element of a tractable family of distributions. Of key interest in statistics and machine learning is Gaussian VI, which approximates by minimizing the Kullback-Leibler (KL) divergence to over the space of Gaussians. In this work, we develop the (Stochastic) Forward-Backward Gaussian Variational Inference (FB-GVI) algorithm to solve Gaussian VI. Our approach exploits the composite structure of the KL divergence, which can be written as the sum of a smooth term (the potential) and a non-smooth term (the entropy) over the Bures-Wasserstein (BW) space of Gaussians endowed with the Wasserstein distance. For our proposed algorithm, we obtain state-of-the-art convergence guarantees when is log-smooth and log-concave, as well as the first convergence guarantees to first-order stationary solutions when is only log-smooth.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers24
- Provable convergence guarantees for black-box variational inferenceJustin Domke, Robert M. Gower, Guillaume GarrigosNeurIPS 2023 · 35 citations
- On the Convergence of Black-Box Variational InferenceKyurae Kim, Jisu Oh, Kaiwen Wu, Yi-An Ma et al.NeurIPS 2023 · 27 citations
- Towards Understanding the Dynamics of Gaussian-Stein Variational Gradient DescentTianle Liu, Promit Ghosal, Krishnakumar Balasubramanian, Natesh S. PillaiNeurIPS 2023 · 19 citations
- Mirror and Preconditioned Gradient Descent in Wasserstein SpaceClément Bonet, Théo Uscidda, Adam David, Pierre-Cyril Aubin-Frankowski et al.NeurIPS 2024 · 19 citations
- Batch and match: black-box variational inference with a score-based divergenceDiana Cai, Chirag Modi, Loucas Pillaud-Vivien, Charles Margossian et al.ICML 2024 · 18 citations
Builds on12
- Variational inference via Wasserstein gradient flowsMarc Lambert, Sinho Chewi, Francis R. Bach, Silvère Bonnabel et al.NeurIPS 2022 · 123 citations
- A Non-Asymptotic Analysis for Stein Variational Gradient DescentAnna Korba, Adil Salim, Michael Arbel, Giulia Luise et al.NeurIPS 2020 · 102 citations
- SVGD as a kernelized Wasserstein gradient flow of the chi-squared divergenceSinho Chewi, Thibaut Le Gouic, Chen Lu, Tyler Maunu et al.NeurIPS 2020 · 92 citations
- Efficient constrained sampling via the mirror-Langevin algorithmKwangjun Ahn, Sinho ChewiNeurIPS 2021 · 77 citations
- The Wasserstein Proximal Gradient AlgorithmAdil Salim, Anna Korba, Giulia LuiseNeurIPS 2020 · 74 citations
Related papers
- Stochastic variance-reduced Gaussian variational inference on the Bures-Wasserstein manifoldHoang Phuc Hau Luu, Hanlin Yu, Bernardo Williams, Marcelo Hartmann et al.ICLR 2025
- Stochastic Gradient Variational Inference with Price's Gradient Estimator from Bures-Wasserstein to Parameter SpaceKyurae Kim, Qiang Fu, Yian Ma, Jacob Gardner et al.ICML 2026
- Variational Inference with Mixtures of Isotropic GaussiansMarguerite Petit-Talamon, Marc Lambert, Anna KorbaNeurIPS 2025 · 7 citations
- Variational inference via Gaussian interacting particles in the Bures-Wasserstein geometryGiacomo Borghi, Jose CarrilloICML 2026 · 2 citations
- Theoretical Guarantees for Variational Inference with Fixed-Variance Mixture of GaussiansTom Huix, Anna Korba, Alain Oliviero Durmus, Eric MoulinesICML 2024 · 12 citations
