Achieving Lossless Gradient Sparsification via Mapping to Alternative Space in Federated Learning
Do-Yeon Kim, Dong-Jun Han, Jun Seo, Jaekyun Moon
Abstract
Handling the substantial communication burden in federated learning (FL) still remains a significant challenge. Although recent studies have attempted to compress the local gradients to address this issue, they typically perform compression only within the original parameter space, which may potentially limit the fundamental compression rate of the gradient. In this paper, instead of restricting our scope to a fixed traditional space, we consider an alternative space that provides an improved compressibility of the gradient. To this end, we utilize the structures of input activation and output gradient in designing our mapping function to a new space, which enables lossless gradient sparsification, i.e., mapping the gradient to our new space induces a greater number of nearzero elements without any loss of information. In light of this attribute, employing sparsificationbased compressors in our new space allows for more aggressive compression with minimal information loss than the baselines. More surprisingly, our model even reaches higher accuracies than the full gradient uploading strategy in some cases, an extra benefit for utilizing the new space. We also theoretically confirm that our approach does not alter the existing, best known convergence rate of FL thanks to the orthogonal transformation properties of our mapping.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 1494c625-5a24-4d2f-ae27-eac8c6a35a10Cited by top-tier papers1
Ask how each one uses itBuilds on17
- On the Convergence of FedAvg on Non-IID DataXiang Li, Kaixuan Huang, Wenhao Yang, Shusen Wang et al.ICLR 2020 · 2,930 citations
- Adaptive Federated OptimizationSashank J. Reddi, Zachary Charles, Manzil Zaheer, Zachary Garrett et al.ICLR 2021 · 1,917 citations
- FetchSGD: Communication-Efficient Federated Learning with SketchingDaniel Rothchild, Ashwinee Panda, Enayat Ullah, Nikita Ivkin et al.ICML 2020 · 425 citations
- Achieving Linear Speedup with Partial Worker Participation in Non-IID Federated LearningHaibo Yang, Minghong Fang, Jia LiuICLR 2021 · 310 citations
- Learn from Others and Be Yourself in Heterogeneous Federated LearningWenke Huang, Mang Ye, Bo DuCVPR 2022 · 254 citations
Related papers
- FedTC: Enabling Communication-Efficient Federated Learning via Transform CodingYixuan Guan, Xuefeng Liu, Jianwei Niu, Tao RenINFOCOM 2024 · 4 citations
- Linear Convergence in Federated Learning: Tackling Client Heterogeneity and Sparse GradientsAritra Mitra, Rayana H. Jaafar, George J. Pappas, Hamed HassaniNeurIPS 2021 · 193 citations
- γ-FedHT: Stepsize-Aware Hard-Threshold Gradient Compression in Federated LearningRongwei Lu, Yutong Jiang, Jinrui Zhang, Chunyang Li et al.INFOCOM 2025 · 2 citations
- On the Discrepancy between the Theoretical Analysis and Practical Implementations of Compressed Communication for Distributed Deep LearningAritra Dutta, El Houcine Bergou, Ahmed M. Abdelmoniem, Chen-Yu Ho et al.AAAI 2020
- SVDFed: Enabling Communication-Efficient Federated Learning via Singular-Value-DecompositionHaolin Wang, Xuefeng Liu, Jianwei Niu, Shaojie TangINFOCOM 2023 · 11 citations
