Asymptotically Optimal Exact Minibatch Metropolis-Hastings
Ruqi Zhang, A. Feder Cooper, Christopher De Sa
Abstract
Metropolis-Hastings (MH) is a commonly-used MCMC algorithm, but it can be intractable on large datasets due to requiring computations over the whole dataset. In this paper, we study minibatch MH methods, which instead use subsamples to enable scaling. We observe that most existing minibatch MH methods are inexact (i.e. they may change the target distribution), and show that this inexactness can cause arbitrarily large errors in inference. We propose a new exact minibatch MH method, TunaMH, which exposes a tunable trade-off between its batch size and its theoretically guaranteed convergence rate. We prove a lower bound on the batch size that any minibatch MH method must use to retain exactness while guaranteeing fast convergence-the first such bound for minibatch MH-and show TunaMH is asymptotically optimal in terms of the batch size. Empirically, we show TunaMH outperforms other exact minibatch MH methods on robust linear regression, truncated Gaussian mixtures, and logistic regression.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 572067be-d57f-4cf1-a5ac-a239ad4bbe08Cited by top-tier papers5
- Surrogate Likelihoods for Variational Annealed Importance SamplingMartin Jankowiak, Du PhanICML 2022 · 14 citations
- Training Bayesian Neural Networks with Sparse Subspace Variational InferenceJunbo Li, Zichen Miao, Qiang Qiu, Ruqi ZhangICLR 2024 · 12 citations
- Entropy-MCMC: Sampling from Flat Basins with EaseBolian Li, Ruqi ZhangICLR 2024 · 7 citations
- DP-Fast MH: Private, Fast, and Accurate Metropolis-Hastings for Large-Scale Bayesian InferenceWanrong Zhang, Ruqi ZhangICML 2023 · 4 citations
- Markov Chain Monte Carlo without Evaluating the Target: an Auxiliary Variable ApproachWei Yuan, Guanyang WangICML 2026 · 3 citations
Related papers
- Spectral Subsampling MCMC for Stationary Time SeriesRobert Salomone, Matias Quiroz, Robert Kohn, Mattias Villani et al.ICML 2020 · 14 citations
- An Even More Optimal Stochastic Optimization Algorithm: Minibatching and Interpolation LearningBlake E. Woodworth, Nathan SrebroNeurIPS 2021 · 22 citations
- A Hybrid Stochastic Gradient Hamiltonian Monte Carlo MethodChao Zhang, Zhijian Li, Zebang Shen, Jiahao Xie et al.AAAI 2021 · 3 citations
- Differentiable Annealed Importance Sampling and the Perils of Gradient NoiseGuodong Zhang, Kyle Hsu, Jianing Li, Chelsea Finn et al.NeurIPS 2021 · 46 citations
- Hamiltonian Monte Carlo Inference of Marginalized Linear Mixed-Effects ModelsJinlin Lai, Justin Domke, Daniel R. SheldonNeurIPS 2024 · 2 citations
