Hierarchical Federated Learning with Multi-Timescale Gradient Correction
Wenzhi Fang, Dong-Jun Han, Evan Chen, Shiqiang Wang, Christopher G. Brinton
Abstract
While traditional federated learning (FL) typically focuses on a star topology where clients are directly connected to a central server, real-world distributed systems often exhibit hierarchical architectures. Hierarchical FL (HFL) has emerged as a promising solution to bridge this gap, leveraging aggregation points at multiple levels of the system. However, existing algorithms for HFL encounter challenges in dealing with multi-timescale model drift, i.e., model drift occurring across hierarchical levels of data heterogeneity. In this paper, we propose a multi-timescale gradient correction (MTGC) methodology to resolve this issue. Our key idea is to introduce distinct control variables to (i) correct the client gradient towards the group gradient, i.e., to reduce client model drift caused by local updates based on individual datasets, and (ii) correct the group gradient towards the global gradient, i.e., to reduce group model drift caused by FL over clients within the group. We analytically characterize the convergence behavior of MTGC under general non-convex settings, overcoming challenges associated with couplings between correction terms. We show that our convergence bound is immune to the extent of data heterogeneity, confirming the stability of the proposed algorithm against multi-level non-i.i.d. data. Through extensive experiments on various datasets and models, we validate the effectiveness of MTGC in diverse HFL settings. The code for this project is available at https://github.com/wenzhifang/MTGC.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 7e5bdc7e-2fd2-4898-897b-27eb01d76544Cited by top-tier papers6
- FIARSE: Model-Heterogeneous Federated Learning via Importance-Aware Submodel ExtractionFeijie Wu, Xingchen Wang, Yaqing Wang, Tianci Liu et al.NeurIPS 2024 · 47 citations
- Federated Sketching LoRA: A Flexible Framework for Heterogeneous Collaborative Fine-Tuning of LLMsWenzhi Fang, Dong-Jun Han, Liangqi Yuan, Seyyedali Hosseinalipour et al.ICML 2026 · 4 citations
- HFedATM: Hierarchical Federated Domain Generalization via Optimal Transport and Regularized Mean AggregationThinh Nguyen, Trung Phan, Binh T. Nguyen, Khoa D. Doan et al.CVPR 2026
- Unified Analyses for Hierarchical Federated Learning: Topology Selection under Data HeterogeneityZiyi Zhou, Yipeng Li, Xinchen LyuICLR 2026
- Towards Federated RLHF with Aggregated Client Preference for LLMsFeijie Wu, Xiaoze Liu, Haoyu Wang, Xingchen Wang et al.ICLR 2025
Builds on17
- SCAFFOLD: Stochastic Controlled Averaging for Federated LearningSai Praneeth Karimireddy, Satyen Kale, Mehryar Mohri, Sashank J. Reddi et al.ICML 2020 · 3,875 citations
- On the Convergence of FedAvg on Non-IID DataXiang Li, Kaixuan Huang, Wenhao Yang, Shusen Wang et al.ICLR 2020 · 2,930 citations
- Tackling the Objective Inconsistency Problem in Heterogeneous Federated OptimizationJianyu Wang, Qinghua Liu, Hao Liang, Gauri Joshi et al.NeurIPS 2020 · 2,231 citations
- A Unified Theory of Decentralized SGD with Changing Topology and Local UpdatesAnastasia Koloskova, Nicolas Loizou, Sadra Boreiri, Martin Jaggi et al.ICML 2020 · 623 citations
- FedSplit: an algorithmic framework for fast federated optimizationReese Pathak, Martin J. WainwrightNeurIPS 2020 · 217 citations
Related papers
- FedDC: Federated Learning with Non-IID Data via Local Drift Decoupling and CorrectionLiang Gao, Huazhu Fu, Li Li, Yingwen Chen et al.CVPR 2022 · 307 citations
- FedLAGC: Towards High Performance System-Heterogeneous Federated Learning via Layer-Adaptive Submodel Extraction and Gradient CorrectionQing Hu, Tianchi Liao, Shuyi Wu, Lei Yang et al.AAAI 2026
- On the Effectiveness of Partial Variance Reduction in Federated Learning with Heterogeneous DataBo Li, Mikkel N. Schmidt, Tommy S. Alstrøm, Sebastian U. StichCVPR 2023
- Efficient Asynchronous Federated Learning with Prospective Momentum Aggregation and Fine-Grained CorrectionYu Zang, Zhe Xue, Shilong Ou, Lingyang Chu et al.AAAI 2024 · 25 citations
- Heterogeneity-Guided Client Sampling: Towards Fast and Efficient Non-IID Federated LearningHuancheng Chen, Haris VikaloNeurIPS 2024 · 15 citations
