Transforming Gaps into Gains: Bridging Model and Data Heterogeneity in Federated Learning via Knowledge Weak-Aware Zones
Ke Li, Yan Ding, Zhiqin Zhu, Shenhai Zheng
Abstract
Heterogeneous federated learning enables collaborative training across clients under dual heterogeneity of models and data, posing challenges for effective knowledge transfer. Federated mutual learning employs proxy models to bridge cross-model knowledge exchange; however, existing methods remain limited to direct alignment between the outputs of private and proxy models, ignoring the deep discrepancies in representation and decision spaces between them. Such cognitive biases cause knowledge to be transferred only at shallow levels and trigger performance bottlenecks. To address this, this paper proposes FedKWAZ to identify and exploit Knowledge Weak-Aware Zones (KWAZ)—spatial zones of deep knowledge misalignment between private and proxy models, further refined into Semantic Weak-Aware Zones and Decision Weak-Aware Zones, which characterize cognitive misalignments in representation and decision spaces as focal targets for enhanced bidirectional distillation. FedKWAZ designs a Hierarchical Adaptive Patch Mixing (HAPM) mechanism to generate multiple mixed samples and employs a Knowledge Discrepancy Perceptron (KDP) to select the samples exhibiting the largest representation and decision discrepancies, thereby mining critical KWAZ. These modules are integrated into a two-stage mutual learning framework, achieving global class-level representation-decision consistency alignment and local KWAZ-guided refinement, structurally bridging cognitive biases across heterogeneous mutual learning models. Experimental results on multiple datasets and model configurations demonstrate the superior performance of FedKWAZ.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 2777a197-37bb-48ce-bf6d-1cd163c8f333Builds on21
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- Exploiting Shared Representations for Personalized Federated LearningLiam Collins, Hamed Hassani, Aryan Mokhtari, Sanjay ShakkottaiICML 2021 · 1,081 citations
- Data-Free Knowledge Distillation for Heterogeneous Federated LearningZhuangdi Zhu, Junyuan Hong, Jiayu ZhouICML 2021 · 957 citations
- FedProto: Federated Prototype Learning across Heterogeneous ClientsYue Tan, Guodong Long, Lu Liu, Tianyi Zhou et al.AAAI 2022 · 851 citations
- FedALA: Adaptive Local Aggregation for Personalized Federated LearningJianqing Zhang, Yang Hua, Hao Wang, Tao Song et al.AAAI 2023 · 445 citations
Related papers
- DKDR: Dynamic Knowledge Distillation for Reliability in Federated LearningYueyang Yuan, Wenke Huang, Frank Wan, Kaiqi Guan et al.NeurIPS 2025 · 1 citation
- A Hierarchical Knowledge Transfer Framework for Heterogeneous Federated LearningYongheng Deng, Ju Ren, Cheng Tang, Feng Lyu et al.INFOCOM 2023 · 37 citations
- FedCD: Towards Consolidated Distillation for Heterogeneous Federated LearningYichen Li, Hang Su, Huifa Li, Haolin Yang et al.AAAI 2026
- Feature Distillation is the Better Choice for Model-Heterogeneous Federated LearningYichen Li, Xiuying Wang, Wenchao Xu, Haozhao Wang et al.NeurIPS 2025 · 6 citations
- FedCDWA: Decoupled Federated Prototype Distillation with Hierarchical Wasserstein AggregationZhenshen Liu, Kai Fan, Wenjie Li, Kuan Zhang et al.ICML 2026
