Aggregating Capacity in FL through Successive Layer Training for Computationally-Constrained Devices
Kilian Pfeiffer, Ramin Khalili, Jörg Henkel
Abstract
Federated learning (FL) is usually performed on resource-constrained edge devices, e.g., with limited memory for the computation. If the required memory to train a model exceeds this limit, the device will be excluded from the training. This can lead to a lower accuracy as valuable data and computation resources are excluded from training, also causing bias and unfairness. The FL training process should be adjusted to such constraints. The state-of-the-art techniques propose training subsets of the FL model at constrained devices, reducing their resource requirements for training. However, these techniques largely limit the co-adaptation among parameters of the model and are highly inefficient, as we show: it is actually better to train a smaller (less accurate) model by the system where all the devices can train the model end-to-end than applying such techniques. We propose a new method that enables successive freezing and training of the parameters of the FL model at devices, reducing the training's resource requirements at the devices while still allowing enough co-adaptation between parameters. We show through extensive experimental evaluation that our technique greatly improves the accuracy of the trained model (by 52.4 p.p.) compared with the state of the art, efficiently aggregating the computation capacity available on distributed devices.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers2
- FedSPU: Personalized Federated Learning for Resource-Constrained Devices with Stochastic Parameter UpdateZiru Niu, Hai Dong, A. K. QinAAAI 2025 · 7 citations
- Bridging Generalization Gap of Heterogeneous Federated Clients Using Generative ModelsZiru Niu, Hai Dong, A. K. QinICLR 2026 · 3 citations
Builds on11
- FetchSGD: Communication-Efficient Federated Learning with SketchingDaniel Rothchild, Ashwinee Panda, Enayat Ullah, Nikita Ivkin et al.ICML 2020 · 425 citations
- FjORD: Fair and Accurate Federated Learning under heterogeneous targets with Ordered DropoutSamuel Horváth, Stefanos Laskaridis, Mário Almeida, Ilias Leontiadis et al.NeurIPS 2021 · 390 citations
- TiFL: A Tier-based Federated Learning SystemZheng Chai, Ahsan Ali, Syed Zawad, Stacey Truex et al.HPDC 2020 · 330 citations
- FedRolex: Model-Heterogeneous Federated Learning with Rolling Sub-Model ExtractionSamiul Alam, Luyang Liu, Ming Yan, Mi ZhangNeurIPS 2022 · 261 citations
- HeteroFL: Computation and Communication Efficient Federated Learning for Heterogeneous ClientsEnmao Diao, Jie Ding, Vahid TarokhICLR 2021 · 179 citations
Related papers
- Breaking the Memory Wall for Heterogeneous Federated Learning via Progressive TrainingYebo Wu, Li Li, Cheng-Zhong XuKDD 2025 · 2 citations
- REFL: Resource-Efficient Federated LearningAhmed M. Abdelmoniem, Atal Narayan Sahu, Marco Canini, Suhaib A. FahmyEuroSys 2023 · 86 citations
- AOCC-FL: Federated Learning with Aligned Overlapping via Calibrated CompensationHaozhao Wang, Wenchao Xu, Yunfeng Fan, Ruixuan Li et al.INFOCOM 2023 · 8 citations
- Beyond End-to-End: Dynamic Chain Optimization for Private LLM Adaptation on the EdgeYebo Wu, Jingguang Li, Chunlin Tian, KaHou Tam et al.ACL 2026 · 1 citation
- ZeroFL: Efficient On-Device Training for Federated Learning with Local SparsityXinchi Qiu, Javier Fernández-Marqués, Pedro P. B. de Gusmao, Yan Gao et al.ICLR 2022 · 87 citations
