Revisiting the Effects of Stochasticity for Hamiltonian Samplers
Giulio Franzese, Dimitrios Milios, Maurizio Filippone, Pietro Michiardi
Abstract
We revisit the theoretical properties of Hamiltonian stochastic differential equations (SDES) for Bayesian posterior sampling, and we study the two types of errors that arise from numerical SDE simulation: the discretization error and the error due to noisy gradient estimates in the context of data subsampling. Our main result is a novel analysis for the effect of mini-batches through the lens of differential operator splitting, revising previous literature results. The stochastic component of a Hamiltonian SDE is decoupled from the gradient noise, for which we make no normality assumptions. This leads to the identification of a convergence bottleneck: when considering mini-batches, the best achievable error rate is , with being the integrator step size. Our theoretical results are supported by an empirical study on a variety of regression and classification tasks for Bayesian neural networks.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext f84d6d6f-c8ef-4020-a6a8-1cd0110755abCited by top-tier papers1
Ask how each one uses itBuilds on5
- How Good is the Bayes Posterior in Deep Neural Networks Really?Florian Wenzel, Kevin Roth, Bastiaan S. Veeling, Jakub Swiatkowski et al.ICML 2020 · 409 citations
- Finite Versus Infinite Neural Networks: an Empirical StudyJaehoon Lee, Samuel S. Schoenholz, Jeffrey Pennington, Ben Adlam et al.NeurIPS 2020 · 245 citations
- Conformal Symplectic and Relativistic OptimizationGuilherme França, Jeremias Sulam, Daniel P. Robinson, René VidalNeurIPS 2020 · 81 citations
- On the Convergence of Hamiltonian Monte Carlo with Stochastic GradientsDifan Zou, Quanquan GuICML 2021 · 20 citations
- Accelerating the diffusion-based ensemble sampling by non-reversible dynamicsFutoshi Futami, Issei Sato, Masashi SugiyamaICML 2020 · 18 citations
Related papers
- Differentiable Annealed Importance Sampling and the Perils of Gradient NoiseGuodong Zhang, Kyle Hsu, Jianing Li, Chelsea Finn et al.NeurIPS 2021 · 46 citations
- Robust Stochastic Gradient Posterior Sampling with Lattice Based DiscretisationZier Mensch, Lars Holdijk, Samuel Duffield, Maxwell Aifer et al.ICML 2026 · 1 citation
- Noise and Fluctuation of Finite Learning Rate Stochastic Gradient DescentKangqiao Liu, Liu Ziyin, Masahito UedaICML 2021 · 46 citations
- Can Microcanonical Langevin Dynamics Leverage Mini-Batch Gradient Noise?Emanuel Sommer, Kangning Diao, Jakob Robnik, Uros Seljak et al.ICML 2026 · 6 citations
- Hamiltonian Monte Carlo on ReLU Neural Networks is InefficientVu C. Dinh, Lam Si Tung Ho, Cuong V. NguyenNeurIPS 2024 · 3 citations
