Revisiting the Effects of Stochasticity for Hamiltonian Samplers
Giulio Franzese, Dimitrios Milios, Maurizio Filippone, Pietro Michiardi
摘要
We revisit the theoretical properties of Hamiltonian stochastic differential equations (SDES) for Bayesian posterior sampling, and we study the two types of errors that arise from numerical SDE simulation: the discretization error and the error due to noisy gradient estimates in the context of data subsampling. Our main result is a novel analysis for the effect of mini-batches through the lens of differential operator splitting, revising previous literature results. The stochastic component of a Hamiltonian SDE is decoupled from the gradient noise, for which we make no normality assumptions. This leads to the identification of a convergence bottleneck: when considering mini-batches, the best achievable error rate is , with being the integrator step size. Our theoretical results are supported by an empirical study on a variety of regression and classification tasks for Bayesian neural networks.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper5
- How Good is the Bayes Posterior in Deep Neural Networks Really?Florian Wenzel, Kevin Roth, Bastiaan S. Veeling, Jakub Swiatkowski 等ICML 2020 · 被引用 409 次
- Finite Versus Infinite Neural Networks: an Empirical StudyJaehoon Lee, Samuel S. Schoenholz, Jeffrey Pennington, Ben Adlam 等NeurIPS 2020 · 被引用 245 次
- Conformal Symplectic and Relativistic OptimizationGuilherme França, Jeremias Sulam, Daniel P. Robinson, René VidalNeurIPS 2020 · 被引用 81 次
- On the Convergence of Hamiltonian Monte Carlo with Stochastic GradientsDifan Zou, Quanquan GuICML 2021 · 被引用 20 次
- Accelerating the diffusion-based ensemble sampling by non-reversible dynamicsFutoshi Futami, Issei Sato, Masashi SugiyamaICML 2020 · 被引用 18 次
相关 Paper
- Differentiable Annealed Importance Sampling and the Perils of Gradient NoiseGuodong Zhang, Kyle Hsu, Jianing Li, Chelsea Finn 等NeurIPS 2021 · 被引用 46 次
- Robust Stochastic Gradient Posterior Sampling with Lattice Based DiscretisationZier Mensch, Lars Holdijk, Samuel Duffield, Maxwell Aifer 等ICML 2026 · 被引用 1 次
- Noise and Fluctuation of Finite Learning Rate Stochastic Gradient DescentKangqiao Liu, Liu Ziyin, Masahito UedaICML 2021 · 被引用 46 次
- Can Microcanonical Langevin Dynamics Leverage Mini-Batch Gradient Noise?Emanuel Sommer, Kangning Diao, Jakob Robnik, Uros Seljak 等ICML 2026 · 被引用 6 次
- Hamiltonian Monte Carlo on ReLU Neural Networks is InefficientVu C. Dinh, Lam Si Tung Ho, Cuong V. NguyenNeurIPS 2024 · 被引用 3 次
