Markov Chain Score Ascent: A Unifying Framework of Variational Inference with Markovian Gradients
Kyurae Kim, Jisu Oh, Jacob R. Gardner, Adji Bousso Dieng, Hongseok Kim
Abstract
Minimizing the inclusive Kullback-Leibler (KL) divergence with stochastic gradient descent (SGD) is challenging since its gradient is defined as an integral over the posterior. Recently, multiple methods have been proposed to run SGD with biased gradient estimates obtained from a Markov chain. This paper provides the first non-asymptotic convergence analysis of these methods by establishing their mixing rate and gradient variance. To do this, we demonstrate that these methods-which we collectively refer to as Markov chain score ascent (MCSA) methods-can be cast as special cases of the Markov chain gradient descent framework. Furthermore, by leveraging this new understanding, we develop a novel MCSA scheme, parallel MCSA (pMCSA), that achieves a tighter bound on the gradient variance. We demonstrate that this improved theoretical result translates to superior empirical performance. * K. Kim is currently with the University of Pennsylvania.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext dbe52e02-5f35-4d52-aa97-22566b0afd47Cited by top-tier papers8
- On the Convergence of Black-Box Variational InferenceKyurae Kim, Jisu Oh, Kaiwen Wu, Yi-An Ma et al.NeurIPS 2023 · 27 citations
- Parallel Tempering With a Variational ReferenceNikola Surjanovic, Saifuddin Syed, Alexandre Bouchard-Côté, Trevor CampbellNeurIPS 2022 · 23 citations
- The Collusion of Memory and Nonlinearity in Stochastic Approximation With Constant StepsizeDongyan Lucy Huo, Yixuan Zhang, Yudong Chen, Qiaomin XieNeurIPS 2024 · 9 citations
- Learning from A Single Markovian Trajectory: Optimality and Variance ReductionZhenyu Sun, Ermin WeiNeurIPS 2025 · 2 citations
- A hitchhiker's guide to Poisson gradient estimationMichael Ibrahim, Hanqi Zhao, Eli Sennesh, Zhi Li et al.ICML 2026 · 1 citation
Builds on5
- Markovian Score Climbing: Variational Inference with KL(p||q)Christian A. Naesseth, Fredrik Lindsten, David M. BleiNeurIPS 2020 · 67 citations
- Challenges and Opportunities in High Dimensional Variational InferenceAkash Kumar Dhaka, Alejandro Catalina, Manushi Welandawe, Michael Riis Andersen et al.NeurIPS 2021 · 54 citations
- Non-asymptotic Convergence of Adam-type Reinforcement Learning Algorithms under Markovian SamplingHuaqing Xiong, Tengyu Xu, Yingbin Liang, Wei ZhangAAAI 2021 · 37 citations
- BR-SNIS: Bias Reduced Self-Normalized Importance SamplingGabriel Cardoso, Sergey Samsonov, Achille Thin, Eric Moulines et al.NeurIPS 2022 · 21 citations
- On the difficulty of unbiased alpha divergence minimizationTomas Geffner, Justin DomkeICML 2021 · 20 citations
Related papers
- Stochastic Approximate Gradient Descent via the Langevin AlgorithmYixuan Qiu, Xiao WangAAAI 2020 · 5 citations
- Accelerating the diffusion-based ensemble sampling by non-reversible dynamicsFutoshi Futami, Issei Sato, Masashi SugiyamaICML 2020 · 18 citations
- Black-Box Variational Inference as a Parametric Approximation to Langevin DynamicsMatthew D. Hoffman, Yian MaICML 2020 · 16 citations
- Stochastic Gradient Descent under Markovian Sampling SchemesMathieu EvenICML 2023 · 41 citations
- A Non-Asymptotic Analysis for Stein Variational Gradient DescentAnna Korba, Adil Salim, Michael Arbel, Giulia Luise et al.NeurIPS 2020 · 102 citations
