Stochastic variance-reduced Gaussian variational inference on the Bures-Wasserstein manifold
Hoang Phuc Hau Luu, Hanlin Yu, Bernardo Williams, Marcelo Hartmann, Arto Klami
摘要
Optimization in the Bures-Wasserstein space has been gaining popularity in the machine learning community since it draws connections between variational inference and Wasserstein gradient flows. The variational inference objective function of Kullback-Leibler divergence can be written as the sum of the negative entropy and the potential energy, making forward-backward Euler the method of choice. Notably, the backward step admits a closed-form solution in this case, facilitating the practicality of the scheme. However, the forward step is not exact since the Bures-Wasserstein gradient of the potential energy involves "intractable" expectations. Recent approaches propose using the Monte Carlo method -in practice a single-sample estimator -to approximate these terms, resulting in high variance and poor performance. We propose a novel variance-reduced estimator based on the principle of control variates. We theoretically show that this estimator has a smaller variance than the Monte-Carlo estimator in scenarios of interest. We also prove that variance reduction helps improve the optimization bounds of the current analysis. We empirically demonstrate that the proposed estimator gains order-of-magnitude improvements over previous Bures-Wasserstein methods. Recently, there has been emerging interest in Gaussian VI with a new geometric Riemannian optimization perspective (Lambert et al., 2022; Diao et al., 2023). The family of non-degenerate Gaus-
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper8
- Large-Scale Wasserstein Gradient FlowsPetr Mokrov, Alexander Korotin, Lingxiao Li, Aude Genevay 等NeurIPS 2021 · 被引用 112 次
- The Wasserstein Proximal Gradient AlgorithmAdil Salim, Anna Korba, Giulia LuiseNeurIPS 2020 · 被引用 74 次
- Averaging on the Bures-Wasserstein manifold: dimension-free convergence of gradient descentJason M. Altschuler, Sinho Chewi, Patrik Gerber, Austin J. StrommeNeurIPS 2021 · 被引用 60 次
- Forward-Backward Gaussian Variational Inference via JKO in the Bures-Wasserstein SpaceMichael Ziyang Diao, Krishna Balasubramanian, Sinho Chewi, Adil SalimICML 2023 · 被引用 47 次
- Provable convergence guarantees for black-box variational inferenceJustin Domke, Robert M. Gower, Guillaume GarrigosNeurIPS 2023 · 被引用 35 次
相关 Paper
- Variational inference via Wasserstein gradient flowsMarc Lambert, Sinho Chewi, Francis R. Bach, Silvère Bonnabel 等NeurIPS 2022 · 被引用 123 次
- Sliced Wasserstein Estimation with Control VariatesKhai Nguyen, Nhat HoICLR 2024 · 被引用 16 次
- On the Convergence of Projected Bures-Wasserstein Gradient Descent under Euclidean Strong ConvexityJunyi Fan, Yuxuan Han, Zijian Liu, Jian-Feng Cai 等ICML 2024 · 被引用 2 次
- Approximation Based Variance Reduction for Reparameterization GradientsTomas Geffner, Justin DomkeNeurIPS 2020 · 被引用 13 次
- Variational inference via Gaussian interacting particles in the Bures-Wasserstein geometryGiacomo Borghi, Jose CarrilloICML 2026 · 被引用 2 次
