FedID: Federated Interactive Distillation for Large-Scale Pretraining Language Models
Xinge Ma, Jiangming Liu, Jin Wang, Xuejie Zhang
摘要
The growing concerns and regulations surrounding the protection of user data privacy have necessitated decentralized training paradigms. To this end, federated learning (FL) is widely studied in user-related natural language processing (NLP). However, it suffers from several critical limitations including extensive communication overhead, inability to handle heterogeneity, and vulnerability to white-box inference attacks. Federated distillation (FD) is proposed to alleviate these limitations, but its performance is faded by confirmation bias. To tackle this issue, we propose Federated Interactive Distillation (FedID), which utilizes a small amount of labeled data retained by the server to further rectify the local models during knowledge transfer. Additionally, based on the GLUE benchmark, we develop a benchmarking framework across multiple tasks with diverse data distributions to contribute to the research of FD in NLP community. Experiments show that our proposed Fe-dID framework achieves the best results in homogeneous and heterogeneous federated scenarios. The code for this paper is available at: https://github.com/maxinge8698/FedID .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- Learning to Collaborate: An Orchestrated-Decentralized Framework for Peer-to-Peer LLM FederationInderjeet Singh, Eléonore Vissol-Gaudin, Andikan Otung, Motoyoshi SekiyaAAAI 2026 · 被引用 1 次
- Data-Free Black-Box Federated Learning via Zeroth-Order Gradient EstimationXinge Ma, Jin Wang, Xuejie ZhangAAAI 2025
它引用的顶会 Paper5
- Ensemble Distillation for Robust Model Fusion in Federated LearningTao Lin, Lingjing Kong, Sebastian U. Stich, Martin JaggiNeurIPS 2020 · 被引用 1,615 次
- ELECTRA: Pre-training Text Encoders as Discriminators Rather Than GeneratorsKevin Clark, Minh-Thang Luong, Quoc V. Le, Christopher D. ManningICLR 2020 · 被引用 541 次
- FedED: Federated Learning via Ensemble Distillation for Medical Relation ExtractionDianbo Sui, Yubo Chen, Jun Zhao, Yantao Jia 等EMNLP 2020 · 被引用 126 次
- Federated Model Decomposition with Private Vocabulary for Text ClassificationZhuo Zhang, Xiangjing Hu, Lizhen Qu, Qifan Wang 等EMNLP 2022 · 被引用 5 次
- Meta Pseudo LabelsHieu Pham, Zihang Dai, Qizhe Xie, Quoc V. LeCVPR 2021
相关 Paper
- Ensemble Attention Distillation for Privacy-Preserving Federated LearningXuan Gong, Abhishek Sharma, Srikrishna Karanam, Ziyan Wu 等ICCV 2021 · 被引用 148 次
- Data-Free Knowledge Distillation for Heterogeneous Federated LearningZhuangdi Zhu, Junyuan Hong, Jiayu ZhouICML 2021 · 被引用 957 次
- Performance Optimization of Federated Person Re-identification via Benchmark AnalysisWeiming Zhuang, Yonggang Wen, Xuesen Zhang, Xin Gan 等ACM MM 2020 · 被引用 94 次
- MH-pFLID: Model Heterogeneous personalized Federated Learning via Injection and Distillation for Medical Data AnalysisLuyuan Xie, Manqing Lin, Tianyu Luan, Cong Li 等ICML 2024 · 被引用 21 次
- FEDLEGAL: The First Real-World Federated Learning Benchmark for Legal NLPZhuo Zhang, Xiangjing Hu, Jingyuan Zhang, Yating Zhang 等ACL 2023 · 被引用 18 次
