An Efficient and Accurate Dynamic Sparse Training Framework Based on Parameter-Freezing
Lei Li, Haochen Yang, Jiacheng Guo, Hongkai Yu, Minghai Qin, Tianyun Zhang
摘要
Federated learning is a decentralized machine learning approach that consists of servers and clients. It protects data privacy during model training by keeping the training data locally in each client. However, the requirement for the server and clients to frequently synchronize the parameters of the model brings a heavy burden to the communication links, especially when the model size has grown drastically in recent years. Several methods have been proposed to compress the model size by sparsification to reduce the communication overhead, albeit with significant accuracy degradation. In this work, we propose methods to better trade-off between model accuracy and training efficiency in federated learning. Our first proposed method is a novel sparse mask readjustment rule on the server and the second is a parameter-freezing method during training on the clients. Experimental results show that the model accuracy has significantly improved when combining our proposed methods. For example, compared with the previous state-of-the-art methods with the same total amount of communication cost and computation FLOPs, the accuracy increases on average by 4% and 6% in our methods for CIFAR-10 and CIFAR-100 datasets on ResNet-18, respectively. On the other hand, when targeting the same accuracy, the proposed method can reduce the communication cost by 4-8 times for different datasets with different sparsity levels.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper7
- The Lottery Ticket Hypothesis for Pre-trained BERT NetworksTianlong Chen, Jonathan Frankle, Shiyu Chang, Sijia Liu 等NeurIPS 2020 · 被引用 428 次
- Dynamic Sparse Training: Find Efficient Sparse Network From Scratch With Trainable Masked LayersJunjie Liu, Zhe Xu, Runbin Shi, Ray C. C. Cheung 等ICLR 2020 · 被引用 136 次
- Federated Dynamic Sparse Training: Computing Less, Communicating Less, Yet Learning BetterSameer Bibikar, Haris Vikalo, Zhangyang Wang, Xiaohan ChenAAAI 2022 · 被引用 133 次
- Sparse Random Networks for Communication-Efficient Federated LearningBerivan Isik, Francesco Pase, Deniz Gündüz, Tsachy Weissman 等ICLR 2023 · 被引用 8 次
- DropIT: Dropping Intermediate Tensors for Memory-Efficient DNN TrainingJoya Chen, Kai Xu, Yuhui Wang, Yifei Cheng 等ICLR 2023 · 被引用 2 次
相关 Paper
- SparsyFed: Sparse Adaptive Federated LearningAdriano Guastella, Lorenzo Sani, Alex Iacob, Alessio Mora 等ICLR 2025
- Complement Sparsification: Low-Overhead Model Pruning for Federated LearningXiaopeng Jiang, Cristian BorceaAAAI 2023 · 被引用 36 次
- DisPFL: Towards Communication-Efficient Personalized Federated Learning via Decentralized Sparse TrainingRong Dai, Li Shen, Fengxiang He, Xinmei Tian 等ICML 2022 · 被引用 163 次
- SpaFL: Communication-Efficient Federated Learning With Sparse Models And Low Computational OverheadMinsu Kim, Walid Saad, Mérouane Debbah, Choong Seon HongNeurIPS 2024 · 被引用 30 次
- DAdaQuant: Doubly-adaptive quantization for communication-efficient Federated LearningRobert Hönig, Yiren Zhao, Robert MullinsICML 2022 · 被引用 87 次
