Accelerating the diffusion-based ensemble sampling by non-reversible dynamics
Futoshi Futami, Issei Sato, Masashi Sugiyama
摘要
Posterior distribution approximation is a central task in Bayesian inference. Stochastic gradient Langevin dynamics (SGLD) and its extensions have been practically used and theoretically studied. While SGLD updates a single particle at a time, ensemble methods that update multiple particles simultaneously have been recently gathering attention. Compared with the naive parallelchain SGLD that updates multiple particles independently, ensemble methods update particles with their interactions. Thus, these methods are expected to be more particle-efficient than the naive parallel-chain SGLD because particles can be aware of other particles' behavior through their interactions. Although ensemble methods numerically demonstrated their superior performance, no theoretical guarantee exists to assure such particle-efficiency and it is unclear whether those ensemble methods are really superior to the naive parallel-chain SGLD in the non-asymptotic settings. To cope with this problem, we propose a novel ensemble method that uses a nonreversible Markov chain for the interaction, and we present a non-asymptotic theoretical analysis for our method. Our analysis shows that, for the first time, the interaction causes a faster convergence rate than the naive parallel-chain SGLD in the non-asymptotic setting if the discretization error is appropriately controlled. Numerical experiments show that we can control the discretization error by tuning the interaction appropriately.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- Interacting Contour Stochastic Gradient Langevin DynamicsWei Deng, Siqi Liang, Botao Hao, Guang Lin 等ICLR 2022 · 被引用 13 次
- Breaking Reversibility Accelerates Langevin Dynamics for Non-Convex OptimizationXuefeng Gao, Mert Gürbüzbalaban, Lingjiong ZhuNeurIPS 2020 · 被引用 9 次
- Accelerating Convergence of Replica Exchange Stochastic Gradient MCMC via Variance ReductionWei Deng, Qi Feng, Georgios Karagiannis, Guang Lin 等ICLR 2021 · 被引用 3 次
- Revisiting the Effects of Stochasticity for Hamiltonian SamplersGiulio Franzese, Dimitrios Milios, Maurizio Filippone, Pietro MichiardiICML 2022 · 被引用 3 次
- Alternating Diffusion for Proximal Sampling with Zeroth Order QueriesHirohane Takagi, Atsushi NitandaICLR 2026 · 被引用 1 次
相关 Paper
- Decentralized Langevin Dynamics for Bayesian LearningAnjaly Parayil, He Bai, Jemin George, Prudhvi GurramNeurIPS 2020 · 被引用 10 次
- Robust Stochastic Gradient Posterior Sampling with Lattice Based DiscretisationZier Mensch, Lars Holdijk, Samuel Duffield, Maxwell Aifer 等ICML 2026 · 被引用 1 次
- Aggregated Gradient Langevin DynamicsChao Zhang, Jiahao Xie, Zebang Shen, Peilin Zhao 等AAAI 2020 · 被引用 1 次
- Stochastic Approximate Gradient Descent via the Langevin AlgorithmYixuan Qiu, Xiao WangAAAI 2020 · 被引用 5 次
- Black-Box Variational Inference as a Parametric Approximation to Langevin DynamicsMatthew D. Hoffman, Yian MaICML 2020 · 被引用 16 次
