On the Convergence of Hamiltonian Monte Carlo with Stochastic Gradients
Difan Zou, Quanquan Gu
摘要
Hamiltonian Monte Carlo (HMC), built based on the Hamilton's equation, has been witnessed great success in sampling from high-dimensional posterior distributions. However, it also suffers from computational inefficiency, especially for large training datasets. One common idea to overcome this computational bottleneck is using stochastic gradients, which only queries a mini-batch of training data in each iteration. However, unlike the extensive studies on the convergence analysis of HMC using full gradients, few works focus on establishing the convergence guarantees of stochastic gradient HMC algorithms. In this paper, we propose a general framework for proving the convergence rate of HMC with stochastic gradient estimators, for sampling from strongly log-concave and log-smooth target distributions. We show that the convergence to the target distribution in 2-Wasserstein distance can be guaranteed as long as the stochastic gradient estimator is unbiased and its variance is upper bounded along the algorithm trajectory. We further apply the proposed framework to analyze the convergence rates of HMC with four standard stochastic gradient estimators: mini-batch stochastic gradient (SG), stochastic variance reduced gradient (SVRG), stochastic average gradient (SAGA), and control variate gradient (CVG). Theoretical results explain the inefficiency of mini-batch SG, and suggest that SVRG and SAGA perform better in the tasks with high-precision requirements, while CVG performs better for large dataset. Experiment results verify our theoretical findings.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- A Symmetry-Aware Exploration of Bayesian Neural Network PosteriorsOlivier Laurent, Emanuel Aldea, Gianni FranchiICLR 2024 · 被引用 12 次
- Constrained Exploration via Reflected Replica Exchange Stochastic Gradient Langevin DynamicsHaoyang Zheng, Hengrong Du, Qi Feng, Wei Deng 等ICML 2024 · 被引用 9 次
- Faster Sampling via Stochastic Gradient Proximal SamplerXunpeng Huang, Difan Zou, Hanze Dong, Yian Ma 等ICML 2024 · 被引用 4 次
- Revisiting the Effects of Stochasticity for Hamiltonian SamplersGiulio Franzese, Dimitrios Milios, Maurizio Filippone, Pietro MichiardiICML 2022 · 被引用 3 次
- Accelerating Hamiltonian Monte Carlo via Chebyshev Integration TimeJun-Kun Wang, Andre WibisonoICLR 2023
它引用的顶会 Paper2
- Fractional Underdamped Langevin Dynamics: Retargeting SGD with Momentum under Heavy-Tailed Gradient NoiseUmut Simsekli, Lingjiong Zhu, Yee Whye Teh, Mert GürbüzbalabanICML 2020 · 被引用 58 次
- Non-convex Learning via Replica Exchange Stochastic Gradient MCMCWei Deng, Qi Feng, Liyao Gao, Faming Liang 等ICML 2020 · 被引用 54 次
相关 Paper
- A Hybrid Stochastic Gradient Hamiltonian Monte Carlo MethodChao Zhang, Zhijian Li, Zebang Shen, Jiahao Xie 等AAAI 2021 · 被引用 3 次
- Variance Reduction in Stochastic Particle-Optimization SamplingJianyi Zhang, Yang Zhao, Changyou ChenICML 2020 · 被引用 13 次
- Stochastic Reweighted Gradient DescentAyoub El Hanchi, David A. Stephens, Chris J. MaddisonICML 2022 · 被引用 10 次
- Accelerating Langevin Monte Carlo via Efficient Stochastic Runge-Kutta Methods beyond Log-ConcavityBin Yang, Xiaojie WangICML 2026 · 被引用 1 次
- A Gradient Based Strategy for Hamiltonian Monte Carlo Hyperparameter OptimizationAndrew Campbell, Wenlong Chen, Vincent Stimper, José Miguel Hernández-Lobato 等ICML 2021 · 被引用 20 次
