Efficient Distribution Similarity Identification in Clustered Federated Learning via Principal Angles between Client Data Subspaces
Saeed Vahidian, Mahdi Morafah, Weijia Wang, Vyacheslav Kungurtsev, Chen Chen, Mubarak Shah, Bill Lin
摘要
Clustered federated learning (FL) has been shown to produce promising results by grouping clients into clusters. This is especially effective in scenarios where separate groups of clients have significant differences in the distributions of their local data. Existing clustered FL algorithms are essentially trying to group together clients with similar distributions so that clients in the same cluster can leverage each other's data to better perform federated learning. However, prior clustered FL algorithms attempt to learn these distribution similarities indirectly during training, which can be quite time consuming as many rounds of federated learning may be required until the formation of clusters is stabilized. In this paper, we propose a new approach to federated learning that directly aims to efficiently identify distribution similarities among clients by analyzing the principal angles between the client data subspaces. Each client applies a truncated singular value decomposition (SVD) step on its local data in a single-shot manner to derive a small set of principal vectors, which provides a signature that succinctly captures the main characteristics of the underlying distribution. This small set of principal vectors is provided to the server so that the server can directly identify distribution similarities among the clients to form clusters. This is achieved by comparing the similarities of the principal angles between the client data subspaces spanned by those principal vectors. The approach provides a simple, yet effective clustered FL framework that addresses a broad range of data heterogeneity issues beyond simpler forms of Non-IIDness like label skews. Our clustered FL approach also enables convergence guarantees for non-convex objectives.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper16
- Optimizing the Collaboration Structure in Cross-Silo Federated LearningWenxuan Bao, Haohan Wang, Jun Wu, Jingrui HeICML 2023 · 被引用 39 次
- ParallelSFL: A Novel Split Federated Learning Framework Tackling Heterogeneity IssuesYunming Liao, Yang Xu, Hongli Xu, Zhiwei Yao 等MobiCom 2024 · 被引用 27 次
- Clustered Federated Learning via Gradient-based PartitioningHeasung Kim, Hyeji Kim, Gustavo de VecianaICML 2024 · 被引用 18 次
- Personalized Federated Continual Learning via Multi-Granularity PromptHao Yu, Xin Yang, Xin Gao, Yan Kang 等KDD 2024 · 被引用 12 次
- Data Disparity and Temporal Unavailability Aware Asynchronous Federated Learning for Predictive Maintenance on Transportation FleetsLeonie von Wahl, Niklas Heidenreich, Prasenjit Mitra, Michael Nolting 等AAAI 2024 · 被引用 12 次
它引用的顶会 Paper6
- Practical Secure Aggregation for Privacy-Preserving Machine LearningKallista A. Bonawitz, Vladimir Ivanov, Ben Kreuter, Antonio Marcedone 等CCS 2017 · 被引用 3,936 次
- SCAFFOLD: Stochastic Controlled Averaging for Federated LearningSai Praneeth Karimireddy, Satyen Kale, Mehryar Mohri, Sashank J. Reddi 等ICML 2020 · 被引用 3,875 次
- Tackling the Objective Inconsistency Problem in Heterogeneous Federated OptimizationJianyu Wang, Qinghua Liu, Hao Liang, Gauri Joshi 等NeurIPS 2020 · 被引用 2,231 次
- Personalized Federated Learning with Theoretical Guarantees: A Model-Agnostic Meta-Learning ApproachAlireza Fallah, Aryan Mokhtari, Asuman E. OzdaglarNeurIPS 2020 · 被引用 1,354 次
- Federated Learning on Non-IID Data Silos: An Experimental StudyQinbin Li, Yiqun Diao, Quan Chen, Bingsheng HeICDE 2022 · 被引用 1,110 次
相关 Paper
- FedDAG: Clustered Federated Learning via Global Data and Gradient Integration for Heterogeneous EnvironmentsAnik Pramanik, Murat Kantarcioglu, Vincent Oria, Shantanu SharmaICLR 2026 · 被引用 1 次
- FedCE: Personalized Federated Learning Method based on Clustering EnsemblesLuxin Cai, Naiyue Chen, Yuanzhouhan Cao, Jiahuan He 等ACM MM 2023 · 被引用 27 次
- MSCFL: Model Structure-Aware Clustered Federated Learning for System Heterogeneity and Data DriftYang Xu, Xiaowei Wu, Zifeng Xu, Cheng Zhang 等AAAI 2026
- Heterogeneity-Guided Client Sampling: Towards Fast and Efficient Non-IID Federated LearningHuancheng Chen, Haris VikaloNeurIPS 2024 · 被引用 15 次
- An Efficient Framework for Clustered Federated LearningAvishek Ghosh, Jichan Chung, Dong Yin, Kannan RamchandranNeurIPS 2020 · 被引用 1,329 次
