Non-reversible Parallel Tempering for Deep Posterior Approximation
Wei Deng, Qian Zhang, Qi Feng, Faming Liang, Guang Lin
摘要
Parallel tempering (PT), also known as replica exchange, is the go-to workhorse for simulations of multi-modal distributions. The key to the success of PT is to adopt efficient swap schemes. The popular deterministic even-odd (DEO) scheme exploits the non-reversibility property and has successfully reduced the communication cost from O(P 2 ) to O(P ) given sufficiently many P chains. However, such an innovation largely disappears in big data due to the limited chains and few bias-corrected swaps. To handle this issue, we generalize the DEO scheme to promote non-reversibility and propose a few solutions to tackle the underlying bias caused by the geometric stopping time. Notably, in big data scenarios, we obtain an appealing communication cost O(P log P ) based on the optimal window size. In addition, we also adopt stochastic gradient descent (SGD) with large and constant learning rates as exploration kernels. Such a user-friendly nature enables us to conduct approximation tasks for complex posteriors without much tuning costs.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- Diffusive Gibbs SamplingWenlin Chen, Mingtian Zhang, Brooks Paige, José Miguel Hernández-Lobato 等ICML 2024 · 被引用 21 次
- Constrained Exploration via Reflected Replica Exchange Stochastic Gradient Langevin DynamicsHaoyang Zheng, Hengrong Du, Qi Feng, Wei Deng 等ICML 2024 · 被引用 9 次
它引用的顶会 Paper7
- How Good is the Bayes Posterior in Deep Neural Networks Really?Florian Wenzel, Kevin Roth, Bastiaan S. Veeling, Jakub Swiatkowski 等ICML 2020 · 被引用 409 次
- Cyclical Stochastic Gradient MCMC for Bayesian Deep LearningRuqi Zhang, Chunyuan Li, Jianyi Zhang, Changyou Chen 等ICLR 2020 · 被引用 292 次
- Non-convex Learning via Replica Exchange Stochastic Gradient MCMCWei Deng, Qi Feng, Liyao Gao, Faming Liang 等ICML 2020 · 被引用 54 次
- Tight Nonparametric Convergence Rates for Stochastic Gradient Descent under the Noiseless Linear ModelRaphaël Berthier, Francis R. Bach, Pierre GaillardNeurIPS 2020 · 被引用 49 次
- Interacting Contour Stochastic Gradient Langevin DynamicsWei Deng, Siqi Liang, Botao Hao, Guang Lin 等ICLR 2022 · 被引用 13 次
相关 Paper
- Accelerated Parallel Tempering via Neural TransportsLeo Zhang, Peter Potaptchik, Jiajun He, Yuanqi Du 等ICLR 2026 · 被引用 14 次
- Parallel tempering on optimized pathsSaifuddin Syed, Vittorio Romaniello, Trevor Campbell, Alexandre Bouchard-CôtéICML 2021 · 被引用 28 次
- Test-Time Guidance for Flow-Based Generative Models via Parallel Tempering on Source DistributionsShih-Hsin Wang, Joel Keller, Taos Transue, Drake Brown 等ICML 2026
- Parallel Tempering With a Variational ReferenceNikola Surjanovic, Saifuddin Syed, Alexandre Bouchard-Côté, Trevor CampbellNeurIPS 2022 · 被引用 23 次
- Continuously Tempered PDMP samplersMatthew Sutton, Robert Salomone, Augustin Chevallier, Paul FearnheadNeurIPS 2022 · 被引用 2 次
