Tackling Data Heterogeneity in Federated Learning with Class Prototypes
Yutong Dai, Zeyuan Chen, Junnan Li, Shelby Heinecke, Lichao Sun, Ran Xu
摘要
Introduction Federated learning (FL) [1] is an emerging area that attracts significant interest in the machine learning community due to its capability to allow collaborative learning from decentralized data with privacy protection. However, in FL, clients may have different data distributions, which violates the standard independent and identically distribution (i.i.d) assumption in centralized machine learning. The non-i.i.d phenomenon is known as the data heterogeneity issue and is an acknowledged cause of the performance degradation of the global model [2] . Moreover, from the client's perspective, the global model may not be the best for their tasks. Therefore, personalized federated learning (PFL) emerged as a variant of FL, where personalized models are learned from a combination of the global model and local data to best suit client tasks. While PFL methods address data heterogeneity, class imbalance combined with data heterogeneity remains overlooked. Class imbalance occurs when clients' data consists of different class distributions and the client may not possess samples of a particular class at all. Ideally, the personalized model can perform equally well in all classes that appeared in the local training dataset. For example, medical institutions have different distributions of medical records across diseases [3], and it is crucial that the personalized model can detect local diseases with equal precision. Meanwhile, the currently adopted practice of evaluating the effectiveness of PFL methods can also be biased. Specifically, when evaluating the accuracy, a single balanced testing dataset is split into multiple local testing datasets that match clients' training data distributions. Then each personalized model is tested on the local testing dataset, and the averaged accuracy is reported. However, in the presence of class imbalance, such an evaluation protocol will likely give a biased assessment due to the potential overfitting of the dominant classes. It is tempting to borrow techniques developed for centralized class imbalance learning, like re-sampling or re-weighting the minority classes. However, due to the data heterogeneity in the FL setting, different clients might have different dominant classes and even have different missing classes; hence the direct adoption may not be applicable. Furthermore, re-sampling would require the knowledge of all classes, potentially violating the privacy constraints. Recent works in class imbalanced learning in non-FL settings [4, 5] suggest decoupling the training procedure into the representation learning and classification phases. The representation learning phase aims to build high-quality representations for classification, while the classification phase seeks to balance the decision boundaries among dominant classes and minority classes. Interestingly, FL works such as [6, 7] find that the classifier is the cause of performance drop and suggest that learning strong shared representations can boost performance. Consistent with the findings in prior works, as later shown in Figure 1 , we observe that representations for different classes are uniformly distributed over the representation space and cluster around the class prototype when learned with class-balanced datasets. However, when the training set is class-imbalanced, as is the case for different clients, representations of minority classes overlap with those of majority classes; hence, the representations are of low quality. Motivated by these observations, we propose FedNH (non-parametric head), a novel method that imposes uniformity of the representation space and preserves class semantics to address data heterogeneity with imbalanced classes. We initially distribute class prototypes uniformly in the latent space as an inductive bias to improve the quality of learned representations and smoothly infuse the class semantics into class prototypes to improve the performance of classifiers on local tasks. Our contributions are summarized as follows. • We propose FedNH, a novel method that tackles data heterogeneity with class imbalance by utilizing uniformity and semantics of class prototypes. • We design a new metric to evaluate personalized model performance. This metric is less sensitive to class imbalance and reflects personalized model generalization ability on minority classes.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper25
- No Fear of Classifier Biases: Neural Collapse Inspired Federated Learning with Synthetic and Fixed ClassifierZexi Li, Xinyi Shang, Rui He, Tao Lin 等ICCV 2023 · 被引用 81 次
- Taming Cross-Domain Representation Variance in Federated Prototype Learning with Heterogeneous Data DomainsLei Wang, Jieming Bian, Letian Zhang, Chen Chen 等NeurIPS 2024 · 被引用 36 次
- Federated Learning with Extremely Noisy Clients via Negative DistillationYang Lu, Lin Chen, Yonggang Zhang, Yiliang Zhang 等AAAI 2024 · 被引用 33 次
- FuseFL: One-Shot Federated Learning through the Lens of Causality with Progressive Model FusionZhenheng Tang, Yonggang Zhang, Peijie Dong, Yiu-ming Cheung 等NeurIPS 2024 · 被引用 28 次
- Beyond Federated Prototype Learning: Learnable Semantic Anchors with Hyperspherical Contrast for Domain-Skewed DataLele Fu, Sheng Huang, Yanyi Lai, Tianchi Liao 等AAAI 2025 · 被引用 16 次
它引用的顶会 Paper13
- SCAFFOLD: Stochastic Controlled Averaging for Federated LearningSai Praneeth Karimireddy, Satyen Kale, Mehryar Mohri, Sashank J. Reddi 等ICML 2020 · 被引用 3,875 次
- Ensemble Distillation for Robust Model Fusion in Federated LearningTao Lin, Lingjing Kong, Sebastian U. Stich, Martin JaggiNeurIPS 2020 · 被引用 1,615 次
- Decoupling Representation and Classifier for Long-Tailed RecognitionBingyi Kang, Saining Xie, Marcus Rohrbach, Zhicheng Yan 等ICLR 2020 · 被引用 1,496 次
- Ditto: Fair and Robust Federated Learning Through PersonalizationTian Li, Shengyuan Hu, Ahmad Beirami, Virginia SmithICML 2021 · 被引用 1,313 次
- Exploiting Shared Representations for Personalized Federated LearningLiam Collins, Hamed Hassani, Aryan Mokhtari, Sanjay ShakkottaiICML 2021 · 被引用 1,081 次
相关 Paper
- The Best of Both Worlds: Accurate Global and Personalized Models through Federated Learning with Data-Free Hyper-Knowledge DistillationHuancheng Chen, Chianing Wang, Haris VikaloICLR 2023 · 被引用 11 次
- FedAS: Bridging Inconsistency in Personalized Federated LearningXiyuan Yang, Wenke Huang, Mang YeCVPR 2024 · 被引用 69 次
- Personalized Federated Learning with Feature Alignment and Classifier CollaborationJian Xu, Xinyi Tong, Shao-Lun HuangICLR 2023 · 被引用 35 次
- Interaction-Aware Gaussian Weighting for Clustered Federated LearningAlessandro Licciardi, Davide Leo, Eros Fanì, Barbara Caputo 等ICML 2025
- Class-Wise Federated Averaging for Efficient PersonalizationGyuejeong Lee, Daeyoung ChoiICCV 2025 · 被引用 3 次
