Decentralized Langevin Dynamics for Bayesian Learning
Anjaly Parayil, He Bai, Jemin George, Prudhvi Gurram
摘要
Motivated by decentralized approaches to machine learning, we propose a collaborative Bayesian learning algorithm taking the form of decentralized Langevin dynamics in a non-convex setting. Our analysis show that the initial KL-divergence between the Markov Chain and the target posterior distribution is exponentially decreasing while the error contributions to the overall KL-divergence from the additive noise is decreasing in polynomial time. We further show that the polynomial-term experiences speed-up with number of agents and provide sufficient conditions on the time-varying step-sizes to guarantee convergence to the desired distribution. The performance of the proposed algorithm is evaluated on a wide variety of machine learning tasks. The empirical results show that the performance of individual agents with locally available data is on par with the centralized setting with considerable improvement in the convergence rate.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper1
相关 Paper
- Accelerating the diffusion-based ensemble sampling by non-reversible dynamicsFutoshi Futami, Issei Sato, Masashi SugiyamaICML 2020 · 被引用 18 次
- An Improved Analysis of Gradient Tracking for Decentralized Machine LearningAnastasia Koloskova, Tao Lin, Sebastian U. StichNeurIPS 2021 · 被引用 148 次
- Decentralized TD Tracking with Linear Function Approximation and its Finite-Time AnalysisGang Wang, Songtao Lu, Georgios B. Giannakis, Gerald Tesauro 等NeurIPS 2020 · 被引用 30 次
- A Stochastic Linearized Augmented Lagrangian Method for Decentralized Bilevel OptimizationSongtao Lu, Siliang Zeng, Xiaodong Cui, Mark S. Squillante 等NeurIPS 2022 · 被引用 29 次
- Cross-Gradient Aggregation for Decentralized Learning from Non-IID DataYasaman Esfandiari, Sin Yong Tan, Zhanhong Jiang, Aditya Balu 等ICML 2021 · 被引用 61 次
