Understanding Convergence and Generalization in Federated Learning through Feature Learning Theory
Wei Huang, Ye Shi, Zhongyi Cai, Taiji Suzuki
摘要
Federated Learning (FL) has attracted significant attention as an efficient privacypreserving approach to distributed learning across multiple clients. Despite extensive empirical research and practical applications, a systematic way to theoretically understand the convergence and generalization properties in FL remains limited. This work aims to establish a unified theoretical foundation for understanding FL through feature learning theory. We focus on a scenario where each client employs a two-layer convolutional neural network (CNN) for local training on their own data. Many existing works analyze the convergence of Federated Averaging (FedAvg) under lazy training with linearizing assumptions in weight space. In contrast, our approach tracks the trajectory of signal learning and noise memorization in FL, eliminating the need for these assumptions. We further show that FedAvg can achieve near-zero test error by effectively increasing signal-tonoise ratio (SNR) in feature learning, while local training without communication achieves a large constant test error. This finding highlights the benefits of communication for generalization in FL. Moreover, our theoretical results suggest that a weighted FedAvg method, based on the similarity of input features across clients, can effectively tackle data heterogeneity issues in FL. Experimental results on both synthetic and real-world datasets verify our theoretical conclusions and emphasize the effectiveness of the weighted FedAvg approach. * Corresponding author 2 )), with σ p being the strength of noise. Note that C c=1 µ (c) µ (c) ⊤ /∥µ (c) ∥ 2 2 is introduced to ensure noise vector orthogonal to signal
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper19
- Harmonizing Generalization and Personalization in Federated Prompt LearningTianyu Cui, Hongxia Li, Jingya Wang, Ye ShiICML 2024 · 被引用 31 次
- Federated Learning from Vision-Language Foundation Models: Theoretical Analysis and MethodBikang Pan, Wei Huang, Ye ShiNeurIPS 2024 · 被引用 28 次
- On the Comparison between Multi-modal and Single-modal Contrastive LearningWei Huang, Andi Han, Yongqiang Chen, Yuan Cao 等NeurIPS 2024 · 被引用 26 次
- Provable Benefit of Cutout and CutMix for Feature LearningJunsoo Oh, Chulhee YunNeurIPS 2024 · 被引用 11 次
- Provably Transformers Harness Multi-Concept Word Semantics for Efficient In-Context LearningDake Bu, Wei Huang, Andi Han, Atsushi Nitanda 等NeurIPS 2024 · 被引用 11 次
它引用的顶会 Paper29
- On the Convergence of FedAvg on Non-IID DataXiang Li, Kaixuan Huang, Wenhao Yang, Shusen Wang 等ICLR 2020 · 被引用 2,930 次
- Tackling the Objective Inconsistency Problem in Heterogeneous Federated OptimizationJianyu Wang, Qinghua Liu, Hao Liang, Gauri Joshi 等NeurIPS 2020 · 被引用 2,231 次
- Personalized Federated Learning using HypernetworksAviv Shamsian, Aviv Navon, Ethan Fetaya, Gal ChechikICML 2021 · 被引用 452 次
- Minibatch vs Local SGD for Heterogeneous Distributed LearningBlake E. Woodworth, Kumar Kshitij Patel, Nati SrebroNeurIPS 2020 · 被引用 231 次
- Characterizing Impacts of Heterogeneity in Federated Learning upon Large-Scale Smartphone DataChengxu Yang, Qipeng Wang, Mengwei Xu, Zhenpeng Chen 等WWW 2021 · 被引用 171 次
相关 Paper
- Widening the Network Mitigates the Impact of Data Heterogeneity on FedAvgLike Jian, Dong LiuICML 2025
- FedAvg Converges to Zero Training Loss Linearly for Overparameterized Multi-Layer Neural NetworksBingqing Song, Prashant Khanduri, Xinwei Zhang, Jinfeng Yi 等ICML 2023 · 被引用 10 次
- Understanding Clipping for Federated Learning: Convergence and Client-Level Differential PrivacyXinwei Zhang, Xiangyi Chen, Mingyi Hong, Steven Wu 等ICML 2022 · 被引用 134 次
- Bridging Generalization Gap of Heterogeneous Federated Clients Using Generative ModelsZiru Niu, Hai Dong, A. K. QinICLR 2026 · 被引用 3 次
- Provable Benefits of Local Steps in Heterogeneous Federated Learning for Neural Networks: A Feature Learning PerspectiveYajie Bao, Michael Crawshaw, Mingrui LiuICML 2024 · 被引用 6 次
