Theoretical Guarantees for Variational Inference with Fixed-Variance Mixture of Gaussians
Tom Huix, Anna Korba, Alain Oliviero Durmus, Eric Moulines
Abstract
Variational inference (VI) is a popular approach in Bayesian inference, that looks for the best approximation of the posterior distribution within a parametric family, minimizing a loss that is typically the (reverse) Kullback-Leibler (KL) divergence. Despite its empirical success, the theoretical properties of VI have only received attention recently, and mostly when the parametric family is the one of Gaussians. This work aims to contribute to the theoretical study of VI in the non-Gaussian case by investigating the setting of Mixture of Gaussians with fixed covariance and constant weights. In this view, VI over this specific family can be casted as the minimization of a Mollified relative entropy, i.e. the KL between the convolution (with respect to a Gaussian kernel) of an atomic measure supported on Diracs, and the target distribution. The support of the atomic measure corresponds to the localization of the Gaussian components. Hence, solving variational inference becomes equivalent to optimizing the positions of the Diracs (the particles), which can be done through gradient descent and takes the form of an interacting particle system. We study two sources of error of variational inference in this context when optimizing the mollified relative entropy. The first one is an optimization result, that is a descent lemma establishing that the algorithm decreases the objective at each iteration. The second one is an approximation error, that upper bounds the objective between an optimal finite mixture and the target distribution.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers2
- Stochastic Gradient Variational Inference with Price's Gradient Estimator from Bures-Wasserstein to Parameter SpaceKyurae Kim, Qiang Fu, Yian Ma, Jacob Gardner et al.ICML 2026
- Gaussian Mixture Flow Matching ModelsHansheng Chen, Kai Zhang, Hao Tan, Zexiang Xu et al.ICML 2025
Builds on9
- Variational inference via Wasserstein gradient flowsMarc Lambert, Sinho Chewi, Francis R. Bach, Silvère Bonnabel et al.NeurIPS 2022 · 123 citations
- A Non-Asymptotic Analysis for Stein Variational Gradient DescentAnna Korba, Adil Salim, Michael Arbel, Giulia Luise et al.NeurIPS 2020 · 102 citations
- The Wasserstein Proximal Gradient AlgorithmAdil Salim, Anna Korba, Giulia LuiseNeurIPS 2020 · 74 citations
- Kernel Stein Discrepancy DescentAnna Korba, Pierre-Cyril Aubin-Frankowski, Szymon Majewski, Pierre AblinICML 2021 · 64 citations
- Mirror Descent with Relative Smoothness in Measure Spaces, with application to Sinkhorn and EMPierre-Cyril Aubin-Frankowski, Anna Korba, Flavien LégerNeurIPS 2022 · 61 citations
Related papers
- Variational Inference with Mixtures of Isotropic GaussiansMarguerite Petit-Talamon, Marc Lambert, Anna KorbaNeurIPS 2025 · 7 citations
- Forward-Backward Gaussian Variational Inference via JKO in the Bures-Wasserstein SpaceMichael Ziyang Diao, Krishna Balasubramanian, Sinho Chewi, Adil SalimICML 2023 · 47 citations
- Mixture weights optimisation for Alpha-Divergence Variational InferenceKamélia Daudel, Randal DoucNeurIPS 2021 · 11 citations
- Least squares variational inferenceYvann Le Fay, Nicolas Chopin, Simon BarthelméNeurIPS 2025 · 2 citations
- Sampling with Mollified Interaction Energy DescentLingxiao Li, Qiang Liu, Anna Korba, Mikhail Yurochkin et al.ICLR 2023 · 2 citations
