Lune

ICML2023顶会

EF21-P and Friends: Improved Theoretical Communication Complexity for Distributed Optimization with Bidirectional Compression

Kaja Gruntkowska, Alexander Tyurin, Peter Richtárik

2023年份
35被引次数
11顶会引用

摘要

In this work we focus our attention on distributed optimization problems in the context where the communication time between the server and the workers is non-negligible. We obtain novel methods supporting bidirectional compression (both from the server to the workers and vice versa) that enjoy new state-of-the-art theoretical communication complexity for convex and nonconvex problems. Our bounds are the first that manage to decouple the variance/error coming from the workers-to-server and server-to-workers compression, transforming a multiplicative dependence to an additive one. Moreover, in the convex regime, we obtain the first bounds that match the theoretical communication complexity of gradient descent. Even in this convex regime, our algorithms work with biased gradient estimators, which is non-standard and requires new proof techniques that may be of independent interest. Finally, our theoretical results are corroborated through suitable experiments. Distributed Optimization and Bidirectional Compression In this paper, we consider distributed optimization problems in strongly convex, convex and nonconvex settings. Such problems arise in federated learning (Konečný et al., 2016; McMahan et al., 2017) and in deep learning (Ramesh et al., 2021). In federated learning, a large number of workers/devices/nodes contain local data and communicate with a parameter-server that performs optimization of a function * The work of Kaja Gruntkowska was performed during a Summer research internship in the Optimization and Machine Learning Lab at KAUST led by

问问这篇 Paper

智能体会读完全文。

Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。

可以从这些问题问起

智能体调用

Luneget_paper_fulltext

在 Lune 里问

免费开始,无需绑卡

引用它的顶会 Paper11

问问它们各自怎么用它

它引用的顶会 Paper6

相关 Paper

黄昏的海面,两侧是细线勾勒的悬崖