FLuID: Mitigating Stragglers in Federated Learning using Invariant Dropout
Irene Wang, Prashant J. Nair, Divya Mahajan
摘要
Federated Learning (FL) allows machine learning models to train locally on individual mobile devices, synchronizing model updates via a shared server. This approach safeguards user privacy; however, it generates a heterogeneous training environment due to the varying performance capabilities of devices. As a result, "straggler" devices with lower performance often dictate the overall training time. In this work, we aim to alleviate this performance bottleneck due to stragglers by dynamically load balancing the training across the system. We introduce Invariant Dropout, a method that extracts a sub-model based on the weight update threshold, thereby minimizing potential impacts on accuracy. Building on this dropout technique, we develop an adaptive training framework, Federated Learning using Invariant Dropout (FLuID). FLuID offers a lightweight framework for sub-model extraction to regulate the computational intensity, thereby reducing the load on straggler devices without affecting model quality. Our method leverages neuron updates from non-straggler devices to construct a tailored sub-model for each straggler based on client performance profiling. Unlike prior work, FLuID can dynamically adapt to changes in stragglers as runtime conditions shift. We evaluate FLuID using five real-world mobile clients. The evaluations show that Invariant Dropout maintains baseline model efficiency while alleviating the performance bottleneck of stragglers through a dynamic and lightweight runtime approach. 1
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- Towards Straggler-Resilient Split Federated Learning: An Unbalanced Update ApproachDandan Liang, Jianing Zhang, Evan Chen, Zhe Li 等NeurIPS 2025 · 被引用 8 次
- CATransformers: Carbon Aware Transformers Through Joint Model-Hardware OptimizationIrene Wang, Mostafa Elhoushi, Ekin Sumbul, Samuel Hsia 等NeurIPS 2025 · 被引用 8 次
- Characterizing the Efficiency of Distributed Training: A Power, Performance, and Thermal PerspectiveSeokjin Go, Joongun Park, Spandan More, Hanjiang Wu 等MICRO 2025 · 被引用 7 次
- FLHetBench: Benchmarking Device and State Heterogeneity in Federated LearningJunyuan Zhang, Shuang Zeng, Miao Zhang, Runxi Wang 等CVPR 2024 · 被引用 6 次
它引用的顶会 Paper8
- Group Knowledge Transfer: Federated Learning of Large CNNs at the EdgeChaoyang He, Murali Annavaram, Salman AvestimehrNeurIPS 2020 · 被引用 605 次
- FjORD: Fair and Accurate Federated Learning under heterogeneous targets with Ordered DropoutSamuel Horváth, Stefanos Laskaridis, Mário Almeida, Ilias Leontiadis 等NeurIPS 2021 · 被引用 390 次
- HeteroFL: Computation and Communication Efficient Federated Learning for Heterogeneous ClientsEnmao Diao, Jie Ding, Vahid TarokhICLR 2021 · 被引用 179 次
- FedAT: a high-performance and communication-efficient federated learning system with asynchronous tiersZheng Chai, Yujing Chen, Ali Anwar, Liang Zhao 等SC 2021 · 被引用 140 次
- Resource-Adaptive Federated Learning with All-In-One Neural CompositionYiqun Mei, Pengfei Guo, Mo Zhou, Vishal PatelNeurIPS 2022 · 被引用 62 次
相关 Paper
- FedEL: Federated Elastic Learning for Heterogeneous DevicesLetian Zhang, Bo Chen, Jieming Bian, Lei Wang 等NeurIPS 2025 · 被引用 7 次
- FLUDE: An Efficient Federated Learning Framework with Undependable DevicesShilong Wang, Jianchun Liu, Hongli Xu, Chunming QiaoINFOCOM 2026
- Federated Dynamic Sparse Training: Computing Less, Communicating Less, Yet Learning BetterSameer Bibikar, Haris Vikalo, Zhangyang Wang, Xiaohan ChenAAAI 2022 · 被引用 133 次
- FedSPU: Personalized Federated Learning for Resource-Constrained Devices with Stochastic Parameter UpdateZiru Niu, Hai Dong, A. K. QinAAAI 2025 · 被引用 7 次
- REFL: Resource-Efficient Federated LearningAhmed M. Abdelmoniem, Atal Narayan Sahu, Marco Canini, Suhaib A. FahmyEuroSys 2023 · 被引用 86 次
