Throughput-Optimal Topology Design for Cross-Silo Federated Learning
Othmane Marfoq, Chuan Xu, Giovanni Neglia, Richard Vidal
Abstract
Federated learning usually employs a client-server architecture where an orchestrator iteratively aggregates model updates from remote clients and pushes them back a refined model. This approach may be inefficient in cross-silo settings, as close-by data silos with high-speed access links may exchange information faster than with the orchestrator, and the orchestrator may become a communication bottleneck. In this paper we define the problem of topology design for cross-silo federated learning using the theory of max-plus linear systems to compute the system throughput---number of communication rounds per time unit. We also propose practical algorithms that, under the knowledge of measurable network characteristics, find a topology with the largest throughput or with provable throughput guarantees. In realistic Internet networks with 10 Gbps access links for silos, our algorithms speed up training by a factor 9 and 1.5 in comparison to the master-slave architecture and to state-of-the-art MATCHA, respectively. Speedups are even larger with slower access links.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers9
- Federated Multi-Task Learning under a Mixture of DistributionsOthmane Marfoq, Giovanni Neglia, Aurélien Bellet, Laetitia Kameni et al.NeurIPS 2021 · 415 citations
- Beyond Exponential Graph: Communication-Efficient Topologies for Decentralized Learning via Finite-time ConvergenceYuki Takezawa, Ryoma Sato, Han Bao, Kenta Niwa et al.NeurIPS 2023 · 22 citations
- On the Convergence of Zeroth-Order Federated Tuning for Large Language ModelsZhenqing Ling, Daoyuan Chen, Liuyi Yao, Yaliang Li et al.KDD 2024 · 17 citations
- Self-Driven Entropy Aggregation for Byzantine-Robust Heterogeneous Federated LearningWenke Huang, Zekun Shi, Mang Ye, He Li et al.ICML 2024 · 16 citations
- Parameter Disparities Dissection for Backdoor Defense in Heterogeneous Federated LearningWenke Huang, Mang Ye, Zekun Shi, Guancheng Wan et al.NeurIPS 2024 · 12 citations
Builds on7
- Deep Learning with Differential PrivacyMartín Abadi, Andy Chu, Ian J. Goodfellow, H. Brendan McMahan et al.CCS 2016 · 7,620 citations
- Membership Inference Attacks Against Machine Learning ModelsReza Shokri, Marco Stronati, Congzheng Song, Vitaly ShmatikovS&P 2017 · 5,137 citations
- Practical Secure Aggregation for Privacy-Preserving Machine LearningKallista A. Bonawitz, Vladimir Ivanov, Ben Kreuter, Antonio Marcedone et al.CCS 2017 · 3,936 citations
- SCAFFOLD: Stochastic Controlled Averaging for Federated LearningSai Praneeth Karimireddy, Satyen Kale, Mehryar Mohri, Sashank J. Reddi et al.ICML 2020 · 3,875 citations
- A Unified Theory of Decentralized SGD with Changing Topology and Local UpdatesAnastasia Koloskova, Nicolas Loizou, Sadra Boreiri, Martin Jaggi et al.ICML 2020 · 623 citations
Related papers
- Reducing Training Time in Cross-Silo Federated Learning using Multigraph TopologyTuong Do, Binh X. Nguyen, Vuong Pham, Toan Tran et al.ICCV 2023 · 4 citations
- GeoFL: A Framework for Efficient Geo-Distributed Cross-Device Federated LearningMaolin Gan, Lanpeng Li, Samiul Alam, Li Liu et al.INFOCOM 2025 · 2 citations
- Efficient and Straggler-Resistant Homomorphic Encryption for Heterogeneous Federated LearningNan Yan, Yuqing Li, Jing Chen, Xiong Wang et al.INFOCOM 2024 · 27 citations
- FedAT: a high-performance and communication-efficient federated learning system with asynchronous tiersZheng Chai, Yujing Chen, Ali Anwar, Liang Zhao et al.SC 2021 · 140 citations
- MAS: Towards Resource-Efficient Federated Multiple-Task LearningWeiming Zhuang, Yonggang Wen, Lingjuan Lyu, Shuai ZhangICCV 2023 · 22 citations
