FedAT: a high-performance and communication-efficient federated learning system with asynchronous tiers
Zheng Chai, Yujing Chen, Ali Anwar, Liang Zhao, Yue Cheng, Huzefa Rangwala
摘要
Federated learning (FL) involves training a model over massive distributed devices, while keeping the training data localized and private. This form of collaborative learning exposes new tradeoffs among model convergence speed, model accuracy, balance across clients, and communication cost, with new challenges including:
(1) straggler problem-where clients lag due to data or (computing and network) resource heterogeneity, and (2) communication bottleneck-where a large number of clients communicate their local updates to a central server and bottleneck the server. Many existing FL methods focus on optimizing along only one single dimension of the tradeoff space. Existing solutions use asynchronous model updating or tiering-based, synchronous mechanisms to tackle the straggler problem. However, asynchronous methods can easily create a communication bottleneck, while tiering may introduce biases that favor faster tiers with shorter response latencies.
To address these issues, we present FedAT, a novel Federated learning system with Asynchronous Tiers under Non-i.i.d. training data. FedAT synergistically combines synchronous, intra-tier training and asynchronous, cross-tier training. By bridging the synchronous and asynchronous training through tiering, FedAT minimizes the straggler effect with improved convergence speed and test accuracy. FedAT uses a straggler-aware, weighted aggregation heuristic to steer and balance the training across clients for further accuracy improvement. FedAT compresses uplink and downlink communications using an efficient, polyline-encodingbased compression algorithm, which minimizes the communication cost. Results show that FedAT improves the prediction performance by up to 21.09% and reduces the communication cost by up to 8.5×, compared to state-of-the-art FL methods.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper13
- FederatedScope: A Flexible Federated Learning Platform for HeterogeneityYuexiang Xie, Zhen Wang, Dawei Gao, Daoyuan Chen 等VLDB 2023 · 被引用 120 次
- Asynchronous SGD Beats Minibatch SGD Under Arbitrary DelaysKonstantin Mishchenko, Francis R. Bach, Mathieu Even, Blake E. WoodworthNeurIPS 2022 · 被引用 95 次
- FedVS: Straggler-Resilient and Privacy-Preserving Vertical Federated Learning for Split ModelsSongze Li, Duanyi Yao, Jin LiuICML 2023 · 被引用 49 次
- FLuID: Mitigating Stragglers in Federated Learning using Invariant DropoutIrene Wang, Prashant J. Nair, Divya MahajanNeurIPS 2023 · 被引用 42 次
- FedCompass: Efficient Cross-Silo Federated Learning on Heterogeneous Client Devices Using a Computing Power-Aware SchedulerZilinghan Li, Pranshu Chaturvedi, Shilan He, Han Chen 等ICLR 2024 · 被引用 23 次
它引用的顶会 Paper4
- On the Convergence of FedAvg on Non-IID DataXiang Li, Kaixuan Huang, Wenhao Yang, Shusen Wang 等ICLR 2020 · 被引用 2,930 次
- Fair Resource Allocation in Federated LearningTian Li, Maziar Sanjabi, Ahmad Beirami, Virginia SmithICLR 2020 · 被引用 971 次
- Lessons Learned from the Chameleon TestbedKate Keahey, Jason Anderson, Zhuo Zhen, Pierre Riteau 等USENIX ATC 2020 · 被引用 398 次
- TiFL: A Tier-based Federated Learning SystemZheng Chai, Ahsan Ali, Syed Zawad, Stacey Truex 等HPDC 2020 · 被引用 330 次
相关 Paper
- FedFetch: Faster Federated Learning with Adaptive Downstream PrefetchingQifan Yan, Andrew Liu, Shiqi He, Mathias Lécuyer 等INFOCOM 2025 · 被引用 2 次
- FedEL: Federated Elastic Learning for Heterogeneous DevicesLetian Zhang, Bo Chen, Jieming Bian, Lei Wang 等NeurIPS 2025 · 被引用 7 次
- Federated Learning under Heterogeneous and Correlated Client AvailabilityAngelo Rodio, Francescomaria Faticanti, Othmane Marfoq, Giovanni Neglia 等INFOCOM 2023 · 被引用 27 次
- FedAS: Bridging Inconsistency in Personalized Federated LearningXiyuan Yang, Wenke Huang, Mang YeCVPR 2024 · 被引用 69 次
- No One Idles: Efficient Heterogeneous Federated Learning with Parallel Edge and Server ComputationFeilong Zhang, Xianming Liu, Shiyi Lin, Gang Wu 等ICML 2023 · 被引用 15 次
