Distributed Learning of Fully Connected Neural Networks using Independent Subnet Training
Binhang Yuan, Cameron R. Wolfe, Chen Dun, Yuxin Tang, Anastasios Kyrillidis, Chris Jermaine
摘要
Distributed machine learning (ML) can bring more computational resources to bear than single-machine learning, thus enabling reductions in training time. Distributed learning partitions models and data over many machines, allowing model and dataset sizes beyond the available compute power and memory of a single machine. In practice though, distributed ML is challenging when distribution is mandatory, rather than chosen by the practitioner. In such scenarios, data could unavoidably be separated among workers due to limited memory capacity per worker or even because of data privacy issues. There, existing distributed methods will utterly fail due to dominant transfer costs across workers, or do not even apply. We propose a new approach to distributed fully connected neural network learning, called independent subnet training (IST), to handle these cases. In IST, the original network is decomposed into a set of narrow subnetworks with the same depth. These subnetworks are then trained locally before parameters are exchanged to produce new subnets and the training cycle repeats. Such a naturally "model parallel" approach limits memory usage by storing only a portion of network parameters on each device. Additionally, no requirements exist for sharing data between workers (i.e., subnet training is local and independent) and communication volume and frequency are reduced by decomposing the original network into independent subnets. These properties of IST can cope with issues due to distributed data, slow interconnects, or limited device memory, making IST a suitable approach for cases of mandatory distribution. We show experimentally that IST results in training times that are much lower than common distributed learning approaches.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper8
- Helios: Heterogeneity-Aware Federated Learning with Dynamically Balanced CollaborationZirui Xu, Fuxun Yu, Jinjun Xiong, Xiang ChenDAC 2021 · 被引用 50 次
- RSC: Accelerate Graph Neural Networks Training via Randomized Sparse ComputationsZirui Liu, Shengyuan Chen, Kaixiong Zhou, Daochen Zha 等ICML 2023 · 被引用 24 次
- FedP3: Federated Personalized and Privacy-friendly Network Pruning under Model HeterogeneityKai Yi, Nidham Gazagnadou, Peter Richtárik, Lingjuan LyuICLR 2024 · 被引用 18 次
- Federated Learning Over Images: Vertical Decompositions and Pre-Trained Backbones Are Difficult to BeatErdong Hu, Yuxin Tang, Anastasios Kyrillidis, Chris JermaineICCV 2023 · 被引用 13 次
- Towards a Better Theoretical Understanding of Independent Subnetwork TrainingEgor Shulgin, Peter RichtárikICML 2024 · 被引用 8 次
它引用的顶会 Paper5
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah 等NeurIPS 2020 · 被引用 64,255 次
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- Don't Use Large Mini-batches, Use Local SGDTao Lin, Sebastian U. Stich, Kumar Kshitij Patel, Martin JaggiICLR 2020 · 被引用 462 次
- TiFL: A Tier-based Federated Learning SystemZheng Chai, Ahsan Ali, Syed Zawad, Stacey Truex 等HPDC 2020 · 被引用 330 次
- Inefficiency of K-FAC for Large Batch Size TrainingLinjian Ma, Gabe Montague, Jiayu Ye, Zhewei Yao 等AAAI 2020 · 被引用 24 次
相关 Paper
- Federated Dynamic Sparse Training: Computing Less, Communicating Less, Yet Learning BetterSameer Bibikar, Haris Vikalo, Zhangyang Wang, Xiaohan ChenAAAI 2022 · 被引用 133 次
- Aggregating Capacity in FL through Successive Layer Training for Computationally-Constrained DevicesKilian Pfeiffer, Ramin Khalili, Jörg HenkelNeurIPS 2023 · 被引用 16 次
- Device-Wise Federated Network PruningShangqian Gao, Junyi Li, Zeyu Zhang, Yanfu Zhang 等CVPR 2024
- Workflow Optimization for Parallel Split LearningJoana Tirana, Dimitra Tsigkari, George Iosifidis, Dimitris ChatzopoulosINFOCOM 2024 · 被引用 14 次
- HeteroFL: Computation and Communication Efficient Federated Learning for Heterogeneous ClientsEnmao Diao, Jie Ding, Vahid TarokhICLR 2021 · 被引用 179 次
