Sparse Random Networks for Communication-Efficient Federated Learning
Berivan Isik, Francesco Pase, Deniz Gündüz, Tsachy Weissman, Michele Zorzi
Abstract
One main challenge in federated learning is the large communication cost of exchanging weight updates from clients to the server at each round. While prior work has made great progress in compressing the weight updates through gradient compression methods, we propose a radically different approach that does not update the weights at all. Instead, our method freezes the weights at their initial random values and learns how to sparsify the random network for the best performance. To this end, the clients collaborate in training a stochastic binary mask to find the optimal sparse random network within the original one. At the end of the training, the final model is a sparse network with random weights -or a subnetwork inside the dense random network. We show improvements in accuracy, communication (less than 1 bit per parameter (bpp)), convergence speed, and final model size (less than 1 bpp) over relevant baselines on MNIST, EMNIST, CIFAR-10, and CIFAR-100 datasets, in the low bitrate regime. * First two authors contributed equally. Work done while F.P. was visiting Imperial College London.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext ee791fb9-1978-4c6f-aed0-1874dbe55235Cited by top-tier papers20
- FIARSE: Model-Heterogeneous Federated Learning via Importance-Aware Submodel ExtractionFeijie Wu, Xingchen Wang, Yaqing Wang, Tianci Liu et al.NeurIPS 2024 · 47 citations
- SpaFL: Communication-Efficient Federated Learning With Sparse Models And Low Computational OverheadMinsu Kim, Walid Saad, Mérouane Debbah, Choong Seon HongNeurIPS 2024 · 30 citations
- Exact Optimality of Communication-Privacy-Utility Tradeoffs in Distributed Mean EstimationBerivan Isik, Wei-Ning Chen, Ayfer Özgür, Tsachy Weissman et al.NeurIPS 2023 · 23 citations
- FedBAT: Communication-Efficient Federated Learning via Learnable BinarizationShiwei Li, Wenchao Xu, Haozhao Wang, Xing Tang et al.ICML 2024 · 13 citations
- LASER: Linear Compression in Wireless Distributed OptimizationAshok Vardhan Makkuva, Marco Bondaschi, Thijs Vogels, Martin Jaggi et al.ICML 2024 · 9 citations
Builds on13
- Deep Learning with Differential PrivacyMartín Abadi, Andy Chu, Ian J. Goodfellow, H. Brendan McMahan et al.CCS 2016 · 7,620 citations
- Differentially Private Learning with Adaptive ClippingGalen Andrew, Om Thakkar, Brendan McMahan, Swaroop RamaswamyNeurIPS 2021 · 425 citations
- FetchSGD: Communication-Efficient Federated Learning with SketchingDaniel Rothchild, Ashwinee Panda, Enayat Ullah, Nikita Ivkin et al.ICML 2020 · 425 citations
- DisPFL: Towards Communication-Efficient Personalized Federated Learning via Decentralized Sparse TrainingRong Dai, Li Shen, Fengxiang He, Xinmei Tian et al.ICML 2022 · 163 citations
- The Skellam Mechanism for Differentially Private Federated LearningNaman Agarwal, Peter Kairouz, Ziyu LiuNeurIPS 2021 · 161 citations
Related papers
- An Efficient and Accurate Dynamic Sparse Training Framework Based on Parameter-FreezingLei Li, Haochen Yang, Jiacheng Guo, Hongkai Yu et al.AAAI 2025 · 2 citations
- Masked Random Noise for Communication-Efficient Federated LearningShiwei Li, Yingyi Cheng, Haozhao Wang, Xing Tang et al.ACM MM 2024 · 7 citations
- SparsyFed: Sparse Adaptive Federated LearningAdriano Guastella, Lorenzo Sani, Alex Iacob, Alessio Mora et al.ICLR 2025
- Achieving Lossless Gradient Sparsification via Mapping to Alternative Space in Federated LearningDo-Yeon Kim, Dong-Jun Han, Jun Seo, Jaekyun MoonICML 2024 · 4 citations
- SVDFed: Enabling Communication-Efficient Federated Learning via Singular-Value-DecompositionHaolin Wang, Xuefeng Liu, Jianwei Niu, Shaojie TangINFOCOM 2023 · 11 citations
