Least squares variational inference
Yvann Le Fay, Nicolas Chopin, Simon Barthelmé
摘要
Variational inference seeks the best approximation of a target distribution within a chosen family, where "best" means minimising Kullback-Leibler divergence. When the approximation family is exponential, the optimal approximation satisfies a fixed-point equation. We introduce LSVI (Least Squares Variational Inference), a gradient-free, Monte Carlo-based scheme for the fixed-point recursion, where each iteration boils down to performing ordinary least squares regression on tempered log-target evaluations under the variational approximation. We show that LSVI is equivalent to biased stochastic natural gradient descent and use this to derive convergence rates with respect to the numbers of samples and iterations. When the approximation family is Gaussian, LSVI involves inverting the Fisher information matrix, whose size grows quadratically with dimension d. We exploit the regression formulation to eliminate the need for this inversion, yielding O(d 3 ) complexity in the full-covariance case and O(d) in the mean-field case. Finally, we numerically demonstrate LSVI's performance on various tasks, including logistic regression, discrete variable selection, and Bayesian synthetic likelihood, showing results competitive with state-of-the-art methods, even when gradients are unavailable.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper6
- Mirror Descent with Relative Smoothness in Measure Spaces, with application to Sinkhorn and EMPierre-Cyril Aubin-Frankowski, Anna Korba, Flavien LégerNeurIPS 2022 · 被引用 61 次
- Provable Smoothness Guarantees for Black-Box Variational InferenceJustin DomkeICML 2020 · 被引用 41 次
- Robustness Analysis of Non-Convex Stochastic Gradient Descent using Biased ExpectationsKevin Scaman, Cédric MalherbeNeurIPS 2020 · 被引用 37 次
- Provable convergence guarantees for black-box variational inferenceJustin Domke, Robert M. Gower, Guillaume GarrigosNeurIPS 2023 · 被引用 35 次
- Convergence Rates of Non-Convex Stochastic Gradient Descent Under a Generic Lojasiewicz Condition and Local SmoothnessKevin Scaman, Cédric Malherbe, Ludovic Dos SantosICML 2022 · 被引用 24 次
相关 Paper
- Variational Inference with Mixtures of Isotropic GaussiansMarguerite Petit-Talamon, Marc Lambert, Anna KorbaNeurIPS 2025 · 被引用 7 次
- Bayesian Online Natural Gradient (BONG)Matt Jones, Peter G. Chang, Kevin P. MurphyNeurIPS 2024 · 被引用 20 次
- Theoretical Guarantees for Variational Inference with Fixed-Variance Mixture of GaussiansTom Huix, Anna Korba, Alain Oliviero Durmus, Eric MoulinesICML 2024 · 被引用 12 次
- Stochastic Approximate Gradient Descent via the Langevin AlgorithmYixuan Qiu, Xiao WangAAAI 2020 · 被引用 5 次
- Variational Sparse Inverse Cholesky Approximation for Latent Gaussian Processes via Double Kullback-Leibler MinimizationJian Cao, Myeongjong Kang, Felix Jimenez, Huiyan Sang 等ICML 2023 · 被引用 12 次
