Asymptotically Optimal Exact Minibatch Metropolis-Hastings
Ruqi Zhang, A. Feder Cooper, Christopher De Sa
摘要
Metropolis-Hastings (MH) is a commonly-used MCMC algorithm, but it can be intractable on large datasets due to requiring computations over the whole dataset. In this paper, we study minibatch MH methods, which instead use subsamples to enable scaling. We observe that most existing minibatch MH methods are inexact (i.e. they may change the target distribution), and show that this inexactness can cause arbitrarily large errors in inference. We propose a new exact minibatch MH method, TunaMH, which exposes a tunable trade-off between its batch size and its theoretically guaranteed convergence rate. We prove a lower bound on the batch size that any minibatch MH method must use to retain exactness while guaranteeing fast convergence-the first such bound for minibatch MH-and show TunaMH is asymptotically optimal in terms of the batch size. Empirically, we show TunaMH outperforms other exact minibatch MH methods on robust linear regression, truncated Gaussian mixtures, and logistic regression.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- Surrogate Likelihoods for Variational Annealed Importance SamplingMartin Jankowiak, Du PhanICML 2022 · 被引用 14 次
- Training Bayesian Neural Networks with Sparse Subspace Variational InferenceJunbo Li, Zichen Miao, Qiang Qiu, Ruqi ZhangICLR 2024 · 被引用 12 次
- Entropy-MCMC: Sampling from Flat Basins with EaseBolian Li, Ruqi ZhangICLR 2024 · 被引用 7 次
- DP-Fast MH: Private, Fast, and Accurate Metropolis-Hastings for Large-Scale Bayesian InferenceWanrong Zhang, Ruqi ZhangICML 2023 · 被引用 4 次
- Markov Chain Monte Carlo without Evaluating the Target: an Auxiliary Variable ApproachWei Yuan, Guanyang WangICML 2026 · 被引用 3 次
相关 Paper
- Spectral Subsampling MCMC for Stationary Time SeriesRobert Salomone, Matias Quiroz, Robert Kohn, Mattias Villani 等ICML 2020 · 被引用 14 次
- An Even More Optimal Stochastic Optimization Algorithm: Minibatching and Interpolation LearningBlake E. Woodworth, Nathan SrebroNeurIPS 2021 · 被引用 22 次
- A Hybrid Stochastic Gradient Hamiltonian Monte Carlo MethodChao Zhang, Zhijian Li, Zebang Shen, Jiahao Xie 等AAAI 2021 · 被引用 3 次
- Differentiable Annealed Importance Sampling and the Perils of Gradient NoiseGuodong Zhang, Kyle Hsu, Jianing Li, Chelsea Finn 等NeurIPS 2021 · 被引用 46 次
- Hamiltonian Monte Carlo Inference of Marginalized Linear Mixed-Effects ModelsJinlin Lai, Justin Domke, Daniel R. SheldonNeurIPS 2024 · 被引用 2 次
