Lessons from Generalization Error Analysis of Federated Learning: You May Communicate Less Often!
Milad Sefidgaran, Romain Chor, Abdellatif Zaidi, Yijun Wan
Abstract
We investigate the generalization error of statistical learning models in a Federated Learning (FL) setting. Specifically, we study the evolution of the generalization error with the number of communication rounds between clients and a parameter server (PS), i.e., the effect on the generalization error of how often the clients' local models are aggregated at PS. In our setup, the more the clients communicate with PS the less data they use for local training in each round, such that the amount of training data per client is identical for distinct values of . We establish PAC-Bayes and rate-distortion theoretic bounds on the generalization error that account explicitly for the effect of the number of rounds , in addition to the number of participating devices and individual datasets size . The bounds, which apply to a large class of loss functions and learning algorithms, appear to be the first of their kind for the FL setting. Furthermore, we apply our bounds to FL-type Support Vector Machines (FSVM); and derive (more) explicit bounds in this case. In particular, we show that the generalization bound of FSVM increases with , suggesting that more frequent communication with PS diminishes the generalization power. This implies that the population risk decreases less fast with than does the empirical risk. Moreover, our bound suggests that the generalization error of FSVM decreases faster than that of centralized learning by a factor of . Finally, we provide experimental results obtained using neural networks (ResNet-56) which show evidence that not only may our observations for FSVM hold more generally but also that the population risk may even start to increase beyond some value of .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers4
- Improving Generalization in Federated Learning with Model-Data Mutual Information Regularization: A Posterior Inference ApproachHao Zhang, Chenglin Li, Nuowen Kan, Ziyang Zheng et al.NeurIPS 2024 · 13 citations
- FedFACT: A Provable Framework for Controllable Group-Fairness Calibration in Federated LearningLi Zhang, Zhongxuan Han, Xiaohua Feng, Jiaming Zhang et al.NeurIPS 2025 · 2 citations
- Scaling Law Analysis in Federated Learning: How to Select the Optimal Model Size?Xuanyu Chen, Nan Yang, Shuai Wang, Dong YuanAAAI 2026
- Generalization in Federated Learning: A Conditional Mutual Information FrameworkZiqiao Wang, Cheng Long, Yongyi MaoICML 2025
Builds on8
- SCAFFOLD: Stochastic Controlled Averaging for Federated LearningSai Praneeth Karimireddy, Satyen Kale, Mehryar Mohri, Sashank J. Reddi et al.ICML 2020 · 3,875 citations
- Adaptive Federated OptimizationSashank J. Reddi, Zachary Charles, Manzil Zaheer, Zachary Garrett et al.ICLR 2021 · 1,917 citations
- PAC-Bayes Compression Bounds So Tight That They Can Explain GeneralizationSanae Lotfi, Marc Finzi, Sanyam Kapoor, Andres Potapczynski et al.NeurIPS 2022 · 98 citations
- Federated Composite OptimizationHonglin Yuan, Manzil Zaheer, Sashank J. ReddiICML 2021 · 71 citations
- PAC-Bayes Information BottleneckZifeng Wang, Shao-Lun Huang, Ercan Engin Kuruoglu, Jimeng Sun et al.ICLR 2022 · 42 citations
Related papers
- Rate-Distortion Theoretic Bounds on Generalization Error for Distributed LearningMilad Sefidgaran, Romain Chor, Abdellatif ZaidiNeurIPS 2022 · 24 citations
- Ensemble Distillation for Robust Model Fusion in Federated LearningTao Lin, Lingjing Kong, Sebastian U. Stich, Martin JaggiNeurIPS 2020 · 1,615 citations
- A Unified Analysis of Federated Learning with Arbitrary Client ParticipationShiqiang Wang, Mingyue JiNeurIPS 2022 · 85 citations
- Generalization Bounds for Federated Learning: Fast Rates, Unparticipating Clients and Unbounded LossesXiaolin Hu, Shaojie Li, Yong LiuICLR 2023
- Widening the Network Mitigates the Impact of Data Heterogeneity on FedAvgLike Jian, Dong LiuICML 2025
