Enhancing Federated Learning with In-Cloud Unlabeled Data
Lun Wang, Yang Xu, Hongli Xu, Jianchun Liu, Zhiyuan Wang, Liusheng Huang
摘要
Federated learning (FL) has been widely applied to collaboratively train deep learning (DL) models on massive end devices (i.e., clients). Due to the limited storage capacity and high labeling cost, there are always insufficient data stored and annotated on each client. Conversely, in cloud datacenters, there exist large-scale unlabeled data, which are easy to collect from public access (e.g., social media). Herein, upon the federated semi-supervised learning (FSSL) technology, we propose the Ada-FedSemi system, which leverages both on-device labeled data and in-cloud unlabeled data to boost the performance of DL models. Given the limited communication and massive quantity of the clients, in each training round, we decide to select partial clients to participate in FL, and their local models are aggregated by the parameter server (PS) to produce pseudo-labels for the unlabeled data, which are utilized to enhance the global model. Considering that the number of participating clients and the quality of pseudo-labels will have a significant impact on the training performance (e.g., efficiency and accuracy), we introduce a multi-armed bandit (MAB) based online algorithm to adaptively determine the participating fraction and confidence threshold during federated model training. Extensive experiments on benchmark models and datasets show that, given the same resource budget, the model trained by Ada-FedSemi achieves 3%-14.8 % higher test accuracy than that of the baseline methods. Besides, when achieving the same test accuracy, Ada-FedSemi saves up to 48% training cost, compared with the baselines.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper14
- FixMatch: Simplifying Semi-Supervised Learning with Consistency and ConfidenceKihyuk Sohn, David Berthelot, Nicholas Carlini, Zizhao Zhang 等NeurIPS 2020 · 被引用 5,129 次
- On the Convergence of FedAvg on Non-IID DataXiang Li, Kaixuan Huang, Wenhao Yang, Shusen Wang 等ICLR 2020 · 被引用 2,930 次
- Federated Learning with Matched AveragingHongyi Wang, Mikhail Yurochkin, Yuekai Sun, Dimitris S. Papailiopoulos 等ICLR 2020 · 被引用 1,368 次
- Optimizing Federated Learning on Non-IID Data with Reinforcement LearningHao Wang, Zakhary Kaplan, Di Niu, Baochun LiINFOCOM 2020 · 被引用 1,002 次
- In Defense of Pseudo-Labeling: An Uncertainty-Aware Pseudo-label Selection Framework for Semi-Supervised LearningMamshad Nayeem Rizve, Kevin Duarte, Yogesh S. Rawat, Mubarak ShahICLR 2021 · 被引用 630 次
相关 Paper
- SemiFL: Semi-Supervised Federated Learning for Unlabeled Clients with Alternate TrainingEnmao Diao, Jie Ding, Vahid TarokhNeurIPS 2022 · 被引用 130 次
- SemiDFL: A Semi-Supervised Paradigm for Decentralized Federated LearningXinyang Liu, Pengchao Han, Xuan Li, Bo LiuAAAI 2025 · 被引用 3 次
- Class Balanced Adaptive Pseudo Labeling for Federated Semi-Supervised LearningMing Li, Qingli Li, Yan WangCVPR 2023
- (FL)2: Overcoming Few Labels in Federated Semi-Supervised LearningSeungjoo Lee, Thanh-Long V. Le, Jaemin Shin, Sung-Ju LeeNeurIPS 2024 · 被引用 15 次
- Local or Global: Selective Knowledge Assimilation for Federated Learning with Limited LabelsYae Jee Cho, Gauri Joshi, Dimitrios DimitriadisICCV 2023 · 被引用 11 次
