Robust Stochastic Gradient Posterior Sampling with Lattice Based Discretisation
Zier Mensch, Lars Holdijk, Samuel Duffield, Maxwell Aifer, Patrick Coles, Max Welling, Miranda C. N. Cheng
Abstract
Stochastic-gradient MCMC methods enable scalable Bayesian posterior sampling but often suffer from sensitivity to minibatch size and gradient noise. To address this, we propose Stochastic Gradient Lattice Random Walk (SGLRW), an extension of the Lattice Random Walk discretisation. Unlike conventional Stochastic Gradient Langevin Dynamics (SGLD), SGLRW introduces stochastic noise only through the off-diagonal elements of the update covariance; this yields greater robustness to minibatch size while retaining asymptotic correctness. Furthermore, as a comparison we analyse a natural analogue of SGLD utilising gradient clipping. Experimental validation on Bayesian regression and classification demonstrates that SGLRW remains stable in regimes where SGLD fails, including in the presence of heavy-tailed gradient noise, and matches or improves predictive performance.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 5a13ba59-110e-48af-b3cf-df1df0ffe1e3Builds on8
- Bayesian Deep Learning and a Probabilistic Perspective of GeneralizationAndrew Gordon Wilson, Pavel IzmailovNeurIPS 2020 · 845 citations
- Laplace Redux - Effortless Bayesian Deep LearningErik A. Daxberger, Agustinus Kristiadi, Alexander Immer, Runa Eschenhagen et al.NeurIPS 2021 · 508 citations
- Bayesian Low-rank Adaptation for Large Language ModelsAdam X. Yang, Maxime Robeyns, Xi Wang, Laurence AitchisonICLR 2024 · 111 citations
- Variational Bayesian Last LayersJames Harrison, John Willes, Jasper SnoekICLR 2024 · 75 citations
- BLoB: Bayesian Low-Rank Adaptation by Backpropagation for Large Language ModelsYibin Wang, Haizhou Shi, Ligong Han, Dimitris N. Metaxas et al.NeurIPS 2024 · 63 citations
Related papers
- Accelerating the diffusion-based ensemble sampling by non-reversible dynamicsFutoshi Futami, Issei Sato, Masashi SugiyamaICML 2020 · 18 citations
- Parameter Expanded Stochastic Gradient Markov Chain Monte CarloHyunsu Kim, Giung Nam, Chulhee Yun, Hongseok Yang et al.ICLR 2025
- Can Microcanonical Langevin Dynamics Leverage Mini-Batch Gradient Noise?Emanuel Sommer, Kangning Diao, Jakob Robnik, Uros Seljak et al.ICML 2026 · 6 citations
- Fractional Underdamped Langevin Dynamics: Retargeting SGD with Momentum under Heavy-Tailed Gradient NoiseUmut Simsekli, Lingjiong Zhu, Yee Whye Teh, Mert GürbüzbalabanICML 2020 · 58 citations
- Aggregated Gradient Langevin DynamicsChao Zhang, Jiahao Xie, Zebang Shen, Peilin Zhao et al.AAAI 2020 · 1 citation
