Rethinking the Flat Minima Searching in Federated Learning
Taehwan Lee, Sung Whan Yoon
Abstract
Albeit the success of federated learning (FL) in decentralized training, bolstering the generalization of models by overcoming heterogeneity across clients still remains a huge challenge. To aim at improved generalization of FL, a group of recent works pursues flatter minima of models by employing sharpness-aware minimization in the local training at the client side. However, we observe that the global model, i.e., the aggregated model, does not lie on flat minima of the global objective, even with the effort of flatness searching in local training, which we define as flatness discrepancy. By rethinking and theoretically analyzing flatness searching in FL through the lens of the discrepancy problem, we propose a method called Federated Learning for Global Flatness (FedGF) that explicitly pursues the flatter minima of the global models, leading to the relieved flatness discrepancy and remarkable performance gains in the heterogeneous FL benchmarks.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext cb2ade28-2f24-4527-a2df-9dbdf08973d4Cited by top-tier papers7
- Exploring Vacant Classes in Label-Skewed Federated LearningKuangpu Guo, Yuhe Ding, Jian Liang, Zilei Wang et al.AAAI 2025 · 17 citations
- Rising from Ashes: Generalized Federated Learning via Dynamic Parameter ResetJiahao Wu, Ming Hu, Yanxin Yang, Xiaofei Xie et al.NeurIPS 2025 · 1 citation
- A Flat Minima Perspective on Understanding Augmentations and Model RobustnessWeebum Yoo, Sung Whan YoonAAAI 2026 · 1 citation
- FedRAM: Federated Reweighting and Aggregation for Multi-Task LearningFan Wu, Xinyu Yan, Jiabei Liu, Wei Yang Bryan LimNeurIPS 2025 · 1 citation
- FedScar: Correcting Geometric Bias for Flatness-Consistent Federated LearningJianfeng Lu, YuZhao Xiang, Yue Chen, Gang Li et al.ICML 2026
Builds on10
- SCAFFOLD: Stochastic Controlled Averaging for Federated LearningSai Praneeth Karimireddy, Satyen Kale, Mehryar Mohri, Sashank J. Reddi et al.ICML 2020 · 3,875 citations
- On the Convergence of FedAvg on Non-IID DataXiang Li, Kaixuan Huang, Wenhao Yang, Shusen Wang et al.ICLR 2020 · 2,930 citations
- Adaptive Federated OptimizationSashank J. Reddi, Zachary Charles, Manzil Zaheer, Zachary Garrett et al.ICLR 2021 · 1,917 citations
- Sharpness-aware Minimization for Efficiently Improving GeneralizationPierre Foret, Ariel Kleiner, Hossein Mobahi, Behnam NeyshaburICLR 2021 · 1,861 citations
- Generalized Federated Learning via Sharpness Aware MinimizationZhe Qu, Xingyu Li, Rui Duan, Yao Liu et al.ICML 2022 · 219 citations
Related papers
- Beyond Local Sharpness: Communication-Efficient Global Sharpness-aware Minimization for Federated LearningDebora Caldarola, Pietro Cagnasso, Barbara Caputo, Marco CicconeCVPR 2025
- Flexible Sharpness-Aware Personalized Federated LearningXinda Xing, Qiugang Zhan, Xiurui Xie, Yuning Yang et al.AAAI 2025 · 5 citations
- Improving the Model Consistency of Decentralized Federated LearningYifan Shi, Li Shen, Kang Wei, Yan Sun et al.ICML 2023 · 89 citations
- Locally Estimated Global Perturbations are Better than Local Perturbations for Federated Sharpness-aware MinimizationZiqing Fan, Shengchao Hu, Jiangchao Yao, Gang Niu et al.ICML 2024 · 35 citations
- FedMut: Generalized Federated Learning via Stochastic MutationMing Hu, Yue Cao, Anran Li, Zhiming Li et al.AAAI 2024 · 46 citations
