FedPDG: Prediction Discrepancy–Guided Data Generation for Heterogeneous Federated Learning
Yuqi Wang, Jianwei Niu, Xinghao Wu, Xuefeng Liu, Xin Hao
摘要
One emerging approach to mitigating data heterogeneity in Federated Learning (FL) is to employ diffusion models to generate synthetic data for clients, thereby aligning local data distributions with the global distribution. Prior work has primarily focused on balance-oriented augmentation, which assumes a balanced global class distribution and thus generates samples of rare classes to rebalance each client's local dataset. However, in practice, global data distributions are often inherently imbalanced. Moreover, privacy constraints in FL hinder the server’s ability to accurately estimate the global distribution, rendering balance-oriented augmentation suboptimal. This raises a key, underexplored challenge: How can synthetic data be generated and selected to align local distributions with the true, yet unknown, global distribution? Our key insight is that a model’s performance implicitly reflects the data distribution it has been trained on. Based on this observation, we use the performance discrepancy between local and global models to identify the regions where each client’s local dataset is lacking, and generate corresponding samples for clients. Furthermore, we adapt the diffusion model via preference optimization, enabling it to generate data that better aligns with the true global distribution. Extensive experiments on multiple benchmarks demonstrate that FedPDG outperforms state-of-the-art methods, achieving up to 3.82% improvement.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper21
- Direct Preference Optimization: Your Language Model is Secretly a Reward ModelRafael Rafailov, Archit Sharma, Eric Mitchell, Christopher D. Manning 等NeurIPS 2023 · 被引用 10,924 次
- No Fear of Heterogeneity: Classifier Calibration for Federated Learning with Non-IID DataMi Luo, Fei Chen, Dapeng Hu, Yifan Zhang 等NeurIPS 2021 · 被引用 510 次
- Federated Learning with Label Distribution Skew via Logits CalibrationJie Zhang, Zhiqi Li, Bo Li, Jianghe Xu 等ICML 2022 · 被引用 221 次
- DiffusionSat: A Generative Foundation Model for Satellite ImagerySamar Khanna, Patrick Liu, Linqi Zhou, Chenlin Meng 等ICLR 2024 · 被引用 173 次
- Federated Learning Based on Dynamic RegularizationDurmus Alp Emre Acar, Yue Zhao, Ramon Matas Navarro, Matthew Mattina 等ICLR 2021 · 被引用 114 次
相关 Paper
- Fake It Till Make It: Federated Learning with Consensus-Oriented GenerationRui Ye, Yaxin Du, Zhenyang Ni, Yanfeng Wang 等ICLR 2024 · 被引用 11 次
- Tackling Data Heterogeneity in Federated Learning with Class PrototypesYutong Dai, Zeyuan Chen, Junnan Li, Shelby Heinecke 等AAAI 2023 · 被引用 154 次
- The Best of Both Worlds: Accurate Global and Personalized Models through Federated Learning with Data-Free Hyper-Knowledge DistillationHuancheng Chen, Chianing Wang, Haris VikaloICLR 2023 · 被引用 11 次
- Prior Refinement Is Better: Diffusion-Driven Graph Harmonization for Federated Graph LearningShuman Zhuang, Zhihao Wu, Wei Huang, Luojun Lin 等AAAI 2026
- DYNAFED: Tackling Client Data Heterogeneity with Global DynamicsRenjie Pi, Weizhong Zhang, Yueqi Xie, Jiahui Gao 等CVPR 2023
