LocFedMix-SL: Localize, Federate, and Mix for Improved Scalability, Convergence, and Latency in Split Learning
Seungeun Oh, Jihong Park, Praneeth Vepakomma, Sihun Baek, Ramesh Raskar, Mehdi Bennis, Seong-Lyun Kim
摘要
Split learning (SL) is a promising distributed learning framework that enables to utilize the huge data and parallel computing resources of mobile devices. SL is built upon a model-split architecture, wherein a server stores an upper model segment that is shared by different mobile clients storing its lower model segments. Without exchanging raw data, SL achieves high accuracy and fast convergence by only uploading smashed data from clients and downloading global gradients from the server. Nonetheless, the original implementation of SL sequentially serves multiple clients, incurring high latency with many clients. A parallel implementation of SL has great potential in reducing latency, yet existing parallel SL algorithms resort to compromising scalability and/or convergence speed. Motivated by this, the goal of this article is to develop a scalable parallel SL algorithm with fast convergence and low latency. As a first step, we identify that the fundamental bottleneck of existing parallel SL comes from the model-split and parallel computing architectures, under which the server-client model updates are often imbalanced, and the client models are prone to detach from the server’s model. To fix this problem, by carefully integrating local parallelism, federated learning, and mixup augmentation techniques, we propose a novel parallel SL framework, coined LocFedMix-SL. Simulation results corroborate that LocFedMix-SL achieves improved scalability, convergence speed, and latency, compared to sequential SL as well as the state-of-the-art parallel SL algorithms such as SplitFed and LocSplitFed.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper8
- SplitGP: Achieving Both Generalization and Personalization in Federated LearningDong-Jun Han, Do-Yeon Kim, Minseok Choi, Christopher G. Brinton 等INFOCOM 2023 · 被引用 43 次
- MergeSFL: Split Federated Learning with Feature Merging and Batch Size RegulationYunming Liao, Yang Xu, Hongli Xu, Lun Wang 等ICDE 2024 · 被引用 41 次
- ParallelSFL: A Novel Split Federated Learning Framework Tackling Heterogeneity IssuesYunming Liao, Yang Xu, Hongli Xu, Zhiwei Yao 等MobiCom 2024 · 被引用 27 次
- FedImpro: Measuring and Improving Client Update in Federated LearningZhenheng Tang, Yonggang Zhang, Shaohuai Shi, Xinmei Tian 等ICLR 2024 · 被引用 25 次
- Traceable Federated Continual LearningQiang Wang, Bingyan Liu, Yawen LiCVPR 2024 · 被引用 16 次
它引用的顶会 Paper3
- CutMix: Regularization Strategy to Train Strong Classifiers With Localizable FeaturesSangdoo Yun, Dongyoon Han, Sanghyuk Chun, Seong Joon Oh 等ICCV 2019 · 被引用 5,843 次
- Revisiting Locally Supervised Learning: an Alternative to End-to-end TrainingYulin Wang, Zanlin Ni, Shiji Song, Le Yang 等ICLR 2021 · 被引用 99 次
- Inefficiency of K-FAC for Large Batch Size TrainingLinjian Ma, Gabe Montague, Jiayu Ye, Zhewei Yao 等AAAI 2020 · 被引用 24 次
相关 Paper
- Towards Straggler-Resilient Split Federated Learning: An Unbalanced Update ApproachDandan Liang, Jianing Zhang, Evan Chen, Zhe Li 等NeurIPS 2025 · 被引用 8 次
- FSL-SAGE: Accelerating Federated Split Learning via Smashed Activation Gradient EstimationSrijith Nair, Michael Lin, Peizhong Ju, Amirreza Talebi 等ICML 2025
- Edge-MSL: Split Learning on the Mobile Edge via Multi-Armed BanditsTaejin Kim, Jinhang Zuo, Xiaoxi Zhang, Carlee Joe-WongINFOCOM 2024 · 被引用 6 次
- Convergence Analysis of Split Federated Learning on Heterogeneous DataPengchao Han, Chao Huang, Geng Tian, Ming Tang 等NeurIPS 2024 · 被引用 32 次
- Optimizing Split Federated Learning through Adaptive Pipeline ParallelismZuan Xie, Yang Xu, Yunming Liao, Junhao Cheng 等INFOCOM 2026
