A Simple Data Augmentation for Feature Distribution Skewed Federated Learning
Yunlu Yan, Huazhu Fu, Yuexiang Li, Jinheng Xie, Jun Ma, Guang Yang, Lei Zhu
Abstract
Federated Learning (FL) facilitates collaborative learning among multiple clients in a distributed manner and ensures the security of privacy. However, its performance inevitably degrades with non-Independent and Identically Distributed (non-IID) data. In this paper, we focus on the feature distribution skewed FL scenario, a common non-IID situation in real-world applications where data from different clients exhibit varying underlying distributions. This variation leads to feature shift, which is a key issue of this scenario. While previous works have made notable progress, few pay attention to the data itself, i.e., the root of this issue. The primary goal of this paper is to mitigate feature shift from the perspective of data. To this end, we propose a simple yet remarkably effective input-level data augmentation method, namely FedRDN, which randomly injects the statistical information of the local distribution from the entire federation into the client's data. This is beneficial to improve the generalization of local feature representations, thereby mitigating feature shift. Moreover, our FedRDN is a plug-andplay component, which can be seamlessly integrated into the data augmentation flow with only a few lines of code. Extensive experiments on several datasets show that the performance of various representative FL methods can be further improved by integrating our FedRDN, demonstrating its effectiveness, strong compatibility and generalizability. Code will be released.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 9ed84b8a-fea4-4dea-ba0e-235c8ee494ddCited by top-tier papers8
- FedHarmony: Harmonizing Heterogeneous Label Correlations in Federated Multi-Label LearningZhiqiang Kou, Junxiang Wu, Wenke Huang, Wenwen He et al.CVPR 2026 · 3 citations
- Trustworthy Federated Label Distribution Learning under Annotation Quality DisparityJunxiang Wu, Zhiqiang Kou, Hongwei Zeng, Wenke Huang et al.ICML 2026 · 2 citations
- FedAlign: Differentially Private Distribution Alignment for Non-IID Federated LearningPeng Wu, Jiapeng Zhang, Yingjie Song, Xiong Xiao et al.CVPR 2026
- FedVeer: Self-Adaptive Skew Estimation for Robust Federated LearningYun Xin, Bangqi Pan, Jianfeng Lu, Shuqin Cao et al.ICML 2026
- FedHPro: Federated Hyper-Prototype Learning via Gradient MatchingHuan Wang, Jun Shen, Haoran Li, Zhenyu Yang et al.ICML 2026
Builds on27
- CutMix: Regularization Strategy to Train Strong Classifiers With Localizable FeaturesSangdoo Yun, Dongyoon Han, Sanghyuk Chun, Seong Joon Oh et al.ICCV 2019 · 5,843 citations
- Random Erasing Data AugmentationZhun Zhong, Liang Zheng, Guoliang Kang, Shaozi Li et al.AAAI 2020 · 4,134 citations
- SCAFFOLD: Stochastic Controlled Averaging for Federated LearningSai Praneeth Karimireddy, Satyen Kale, Mehryar Mohri, Sashank J. Reddi et al.ICML 2020 · 3,875 citations
- Moment Matching for Multi-Source Domain AdaptationXingchao Peng, Qinxun Bai, Xide Xia, Zijun Huang et al.ICCV 2019 · 2,239 citations
- Tackling the Objective Inconsistency Problem in Heterogeneous Federated OptimizationJianyu Wang, Qinghua Liu, Hao Liang, Gauri Joshi et al.NeurIPS 2020 · 2,231 citations
Related papers
- FedFA: Federated Feature AugmentationTianfei Zhou, Ender KonukogluICLR 2023 · 8 citations
- FRAug: Tackling Federated Learning with Non-IID Features via Representation AugmentationHaokun Chen, Ahmed Frikha, Denis Krompass, Jindong Gu et al.ICCV 2023 · 44 citations
- FedBN: Federated Learning on Non-IID Features via Local Batch NormalizationXiaoxiao Li, Meirui Jiang, Xiaofei Zhang, Michael Kamp et al.ICLR 2021 · 1,166 citations
- RAFed: Responsive Augmentation and Approximate Update Method for Federated Learning with Non-IID DataYicheng Di, Zhanjie ZhangWWW 2026
- FedMix: Approximation of Mixup under Mean Augmented Federated LearningTehrim Yoon, Sumin Shin, Sung Ju Hwang, Eunho YangICLR 2021 · 226 citations
