Exploiting Similarity for Computation and Communication-Efficient Decentralized Optimization
Yuki Takezawa, Xiaowen Jiang, Anton Rodomanov, Sebastian U. Stich
摘要
Reducing communication complexity is critical for efficient decentralized optimization. The proximal decentralized optimization (PDO) framework is particularly appealing, as methods within this framework can exploit functional similarity among nodes to reduce communication rounds. Specifically, when local functions at different nodes are similar, these methods achieve faster convergence with fewer communication steps. However, existing PDO methods often require highly accurate solutions to subproblems associated with the proximal operator, resulting in significant computational overhead. In this work, we propose the Stabilized Proximal Decentralized Optimization (SPDO) method, which achieves state-of-the-art communication and computational complexities within the PDO framework. Additionally, we refine the analysis of existing PDO methods by relaxing subproblem accuracy requirements and leveraging average functional similarity. Experimental results demonstrate that SPDO significantly outperforms existing methods. Algorithm Reference # Communication # Computation Assumptions Gradient Tracking Koloskova et al. ( 2021 ) Accelerated SONATA Tian et al. ( 2022 ) 1, 2, 3, 4 Accelerated Stabilized-PDO [new] Th. 5, 6 O δ µ(1-ρ) log( L µ ) log( 1 ϵ ) O L µ log( 1 ϵ ) 1, 2, 4
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- Non-Convex Federated Optimization under Cost-Aware Client SelectionXiaowen Jiang, Anton Rodomanov, Sebastian U. StichICLR 2026 · 被引用 1 次
- Understanding MARS: When Scaling Momentum Provably HelpsEgor Shulgin, Tamaz Gadaev, Sarit Khirirat, Peter RichtarikICML 2026
它引用的顶会 Paper8
- A Unified Theory of Decentralized SGD with Changing Topology and Local UpdatesAnastasia Koloskova, Nicolas Loizou, Sadra Boreiri, Martin Jaggi 等ICML 2020 · 被引用 623 次
- An Improved Analysis of Gradient Tracking for Decentralized Machine LearningAnastasia Koloskova, Tao Lin, Sebastian U. StichNeurIPS 2021 · 被引用 148 次
- Bias-Variance Reduced Local SGD for Less Heterogeneous Federated LearningTomoya Murata, Taiji SuzukiICML 2021 · 被引用 61 次
- Optimal Gradient Sliding and its Application to Optimal Distributed Optimization Under SimilarityDmitry Kovalev, Aleksandr Beznosikov, Ekaterina Borodich, Alexander V. Gasnikov 等NeurIPS 2022 · 被引用 26 次
- Federated Optimization with Doubly Regularized Drift CorrectionXiaowen Jiang, Anton Rodomanov, Sebastian U. StichICML 2024 · 被引用 18 次
相关 Paper
- Is Consensus Acceleration Possible in Decentralized Optimization over Slowly Time-Varying Networks?Dmitry Metelev, Alexander Rogozin, Dmitry Kovalev, Alexander V. GasnikovICML 2023 · 被引用 5 次
- DADAO: Decoupled Accelerated Decentralized Asynchronous OptimizationAdel Nabli, Edouard OyallonICML 2023 · 被引用 13 次
- D-SPIDER-SFO: A Decentralized Optimization Algorithm with Faster Convergence Rate for Nonconvex ProblemsTaoxing Pan, Jun Liu, Jie WangAAAI 2020 · 被引用 19 次
- Decentralized Accelerated Proximal Gradient DescentHaishan Ye, Ziang Zhou, Luo Luo, Tong ZhangNeurIPS 2020 · 被引用 37 次
- ProxSkip: Yes! Local Gradient Steps Provably Lead to Communication Acceleration! Finally!Konstantin Mishchenko, Grigory Malinovsky, Sebastian U. Stich, Peter RichtárikICML 2022 · 被引用 200 次
