Explainable Federated Learning via Global–Local Attribution Alignment
Dawood Wasif, Terrence Moore, Chang-Tien Lu, Jin-Hee Cho
Abstract
Federated learning enables on-device training without centralizing data, yet existing systems still struggle to provide explanations that are both locally faithful and globally consistent under strict privacy and bandwidth constraints. Prior approaches either keep explanations siloed across clients, transmit heavy or sensitive artifacts, or replace expressive task models with interpretable surrogates that sacrifice accuracy. We propose xFedAlign, a model-agnostic framework that decouples task optimization in parameter space from explanation coordination in a compact group space. Each client distills a lightweight surrogate to produce private, per-class top-k attribution artifacts, which are robustly aggregated by the server into a Global Explanation Prior that softly aligns client explanations without constraining task learning. Across image, text, and tabular benchmarks with IID and non-IID partitions, xFedAlign matches FedAvg accuracy while consistently reducing explanation drift and improving deletion and insertion AUC relative to Local-XAI, FedAttr-Agg, and Fed-XAI, with only a few kilobytes of additional communication per round. Privacy and robustness evaluations further demonstrate reduced membership inference advantage and increased resistance to attribution poisoning, enabling consistent and trustworthy explanations in federated learning.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 47679b42-556f-41b5-8a28-a0c28cbfb96eBuilds on5
- SCAFFOLD: Stochastic Controlled Averaging for Federated LearningSai Praneeth Karimireddy, Satyen Kale, Mehryar Mohri, Sashank J. Reddi et al.ICML 2020 · 3,875 citations
- Adaptive Federated OptimizationSashank J. Reddi, Zachary Charles, Manzil Zaheer, Zachary Garrett et al.ICLR 2021 · 1,917 citations
- Exploiting Unintended Feature Leakage in Collaborative LearningLuca Melis, Congzheng Song, Emiliano De Cristofaro, Vitaly ShmatikovS&P 2019 · 1,736 citations
- Concept Bottleneck ModelsPang Wei Koh, Thao Nguyen, Yew Siang Tang, Stephen Mussmann et al.ICML 2020 · 1,233 citations
- FedBN: Federated Learning on Non-IID Features via Local Batch NormalizationXiaoxiao Li, Meirui Jiang, Xiaofei Zhang, Michael Kamp et al.ICLR 2021 · 1,166 citations
Related papers
- FedXDS: Leveraging Model Attribution Methods to Counteract Data Heterogeneity in Federated LearningMaximilian Andreas Hoefler, Karsten Müller, Wojciech SamekICCV 2025
- SHERPA: Explainable Robust Algorithms for Privacy-Preserved Federated Learning in Future Networks to Defend Against Data Poisoning AttacksChamara Sandeepa, Bartlomiej Siniarski, Shen Wang, Madhusanka LiyanageS&P 2024 · 19 citations
- FedAlign: Differentially Private Distribution Alignment for Non-IID Federated LearningPeng Wu, Jiapeng Zhang, Yingjie Song, Xiong Xiao et al.CVPR 2026
- FedCDWA: Decoupled Federated Prototype Distillation with Hierarchical Wasserstein AggregationZhenshen Liu, Kai Fan, Wenjie Li, Kuan Zhang et al.ICML 2026
- Minimizing False-Positive Attributions in Explanations of Non-Linear ModelsAnders Gjølbye, Stefan Haufe, Lars Kai HansenNeurIPS 2025 · 3 citations
