Device-Wise Federated Network Pruning
Shangqian Gao, Junyi Li, Zeyu Zhang, Yanfu Zhang, Weidong Cai, Heng Huang
Abstract
Neural network pruning, particularly channel pruning, is a widely used technique for compressing deep learning models to enable their deployment on edge devices with limited resources. Typically, redundant weights or structures are removed to achieve the target resource budget. Although data-driven pruning approaches have proven to be more effective, they cannot be directly applied to federated learning (FL), which has emerged as a popular technique in edge computing applications, because of distributed and confidential datasets. In response to this challenge, we design a new network pruning method for FL. We propose device-wise sub-networks for each device, assuming that the data distribution is similar within each device. These subnetworks are generated through sub-network embeddings and a hypernetwork. To further minimize memory usage and communication costs, we permanently prune the full model to remove weights that are not useful for all devices. During the FL process, we simultaneously train the device-wise subnetworks and the base sub-network to facilitate the pruning process. We then finetune the pruned model with device-wise sub-networks to regain performance. Moreover, we provided the theoretical guarantee of convergence for our method. Our method achieves better performance and resource tradeoff than other well-established network pruning baselines, as demonstrated through extensive experiments on CIFAR-10, CIFAR-100, and TinyImageNet.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext de5e7c97-efb3-4bb1-abbe-51dc9e76d3edCited by top-tier papers3
- DISP-LLM: Dimension-Independent Structural Pruning for Large Language ModelsShangqian Gao, Chi-Heng Lin, Ting Hua, Zheng Tang et al.NeurIPS 2024 · 42 citations
- ALTER: All-in-One Layer Pruning and Temporal Expert Routing for Efficient Diffusion GenerationXiaomeng Yang, Lei Lu, Qihui Fan, Changdi Yang et al.NeurIPS 2025 · 4 citations
- FedCS: Coreset Selection for Federated LearningChenhe Hao, Weiying Xie, Daixun Li, Haonan Qin et al.CVPR 2025
Builds on25
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- A ConvNet for the 2020sZhuang Liu, Hanzi Mao, Chao-Yuan Wu, Christoph Feichtenhofer et al.CVPR 2022 · 6,782 citations
- On the Convergence of FedAvg on Non-IID DataXiang Li, Kaixuan Huang, Wenhao Yang, Shusen Wang et al.ICLR 2020 · 2,930 citations
- Inverting Gradients - How easy is it to break privacy in federated learning?Jonas Geiping, Hartmut Bauermeister, Hannah Dröge, Michael MoellerNeurIPS 2020 · 1,822 citations
- Ditto: Fair and Robust Federated Learning Through PersonalizationTian Li, Shengyuan Hu, Ahmad Beirami, Virginia SmithICML 2021 · 1,313 citations
Related papers
- FedMef: Towards Memory-Efficient Federated Dynamic PruningHong Huang, Weiming Zhuang, Chen Chen, Lingjuan LyuCVPR 2024
- Submodel Extraction for Efficient and Personalized Federated Learning via Optimal TransportZheng Jiang, Nan He, Yiming Chen, Lifeng SunCVPR 2026
- Resilient and Communication Efficient Learning for Heterogeneous Federated SystemsZhuangdi Zhu, Junyuan Hong, Steve Drew, Jiayu ZhouICML 2022 · 46 citations
- FedLPS: Heterogeneous Federated Learning for Multiple Tasks with Local Parameter SharingYongzhe Jia, Xuyun Zhang, Amin Beheshti, Wanchun DouAAAI 2024 · 16 citations
- Hermes: an efficient federated learning framework for heterogeneous mobile clientsAng Li, Jingwei Sun, Pengcheng Li, Yu Pu et al.MobiCom 2021 · 167 citations
