Towards Diverse Device Heterogeneous Federated Learning via Task Arithmetic Knowledge Integration
Mahdi Morafah, Vyacheslav Kungurtsev, Hojin Chang, Chen Chen, Bill Lin
摘要
Federated Learning (FL) has emerged as a promising paradigm for collaborative machine learning, while preserving user data privacy. Despite its potential, standard FL algorithms lack support for diverse heterogeneous device prototypes, which vary significantly in model and dataset sizes-from small IoT devices to large workstations. This limitation is only partially addressed by existing knowledge distillation (KD) techniques, which often fail to transfer knowledge effectively across a broad spectrum of device prototypes with varied capabilities. This failure primarily stems from two issues: the dilution of informative logits from more capable devices by those from less capable ones, and the use of a single integrated logits as the distillation target across all devices, which neglects their individual learning capacities and and the unique contributions of each device. To address these challenges, we introduce TAKFL, a novel KD-based framework that treats the knowledge transfer from each device prototype's ensemble as a separate task, independently distilling each to preserve its unique contributions and avoid dilution. TAKFL also incorporates a KD-based self-regularization technique to mitigate the issues related to the noisy and unsupervised ensemble distillation process. To integrate the separately distilled knowledge, we introduce an adaptive task arithmetic knowledge integration process, allowing each student model to customize the knowledge integration for optimal performance. Additionally, we present theoretical results demonstrating the effectiveness of task arithmetic in transferring knowledge across heterogeneous device prototypes with varying capacities. Comprehensive evaluations of our method across both computer vision (CV) and natural language processing (NLP) tasks demonstrate that TAKFL achieves state-of-the-art results in a variety of datasets and settings, significantly outperforming existing KD-based methods. Our code is released at https://github.com/MMorafah/TAKFL .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper7
- Tackling Intertwined Data and Device Heterogeneities in Federated Learning with Unlimited StalenessHaoming Wang, Wei GaoAAAI 2025 · 被引用 3 次
- FedFree: Breaking Knowledge-sharing Barriers through Layer-wise Alignment in Heterogeneous Federated LearningHaizhou Du, Yiran Xiang, Yiwen Cai, Xiufeng Liu 等NeurIPS 2025 · 被引用 3 次
- C2Prompt: Class-aware Client Knowledge Interaction for Federated Continual LearningKunlun Xu, Yibo Feng, Jiangmeng Li, Yongsheng Qi 等NeurIPS 2025 · 被引用 2 次
- Bi-level Personalization for Federated Foundation Models: A Task-vector Aggregation ApproachYiyuan Yang, Guodong Long, Qinghua Lu, Liming Zhu 等AAAI 2026
- FedARC: Anchor-Guided Residual Compensation for Data and Model Heterogeneous Federated LearningChentao Lu, Xuhao Ren, Dawei xu, Chuan Zhang 等ICML 2026
它引用的顶会 Paper20
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- SCAFFOLD: Stochastic Controlled Averaging for Federated LearningSai Praneeth Karimireddy, Satyen Kale, Mehryar Mohri, Sashank J. Reddi 等ICML 2020 · 被引用 3,875 次
- Tackling the Objective Inconsistency Problem in Heterogeneous Federated OptimizationJianyu Wang, Qinghua Liu, Hao Liang, Gauri Joshi 等NeurIPS 2020 · 被引用 2,231 次
- MobileViT: Light-weight, General-purpose, and Mobile-friendly Vision TransformerSachin Mehta, Mohammad RastegariICLR 2022 · 被引用 2,162 次
相关 Paper
- Ensemble Distillation for Robust Model Fusion in Federated LearningTao Lin, Lingjing Kong, Sebastian U. Stich, Martin JaggiNeurIPS 2020 · 被引用 1,615 次
- FedAKD: Federated Adaptive Knowledge Distillation via Global Knowledge Calibration and DecouplingYingchao Wang, Wenqi Niu, Hanpo HouWWW 2026
- A Hierarchical Knowledge Transfer Framework for Heterogeneous Federated LearningYongheng Deng, Ju Ren, Cheng Tang, Feng Lyu 等INFOCOM 2023 · 被引用 37 次
- FedCD: Towards Consolidated Distillation for Heterogeneous Federated LearningYichen Li, Hang Su, Huifa Li, Haolin Yang 等AAAI 2026
- Towards Understanding Ensemble Distillation in Federated LearningSejun Park, Kihun Hong, Ganguk HwangICML 2023 · 被引用 9 次
