Towards Energy-efficient Federated Learning via INT8-based Training on Mobile DSPs
Jinliang Yuan, Shangguang Wang, Hongyu Li, Daliang Xu, Yuanchun Li, Mengwei Xu, Xuanzhe Liu
摘要
AI is making the Web an even cooler place, but also introduces serious privacy risks due to the extensive user data collection. Federated learning (FL), as a privacy-preserving machine learning paradigm, enables mobile devices to collaboratively learn a shared prediction model while keeping all training data on devices. However, a key obstacle towards practical cross-device FL training is huge energy consumption, especially for lightweight mobile devices. In this work, we perform the first-of-its-kind analysis of improving FL performance through low-precision training with an energy-friendly Digital Signal Processor (DSP) on mobile devices. We first demonstrate that directly integrating the state-of-the-art INT8 (8-bit integer) training algorithm and classic FL protocols will significantly degrade the model accuracy. Moreover, we observe that there are still unavoidable frequent quantization operations on devices that cause extreme load stress on DSP-enabled INT8 training. To address the above challenges, we present Q-FedUpdate, an FL framework that efficiently preserves model accuracy with ultra-low energy consumption. It maintains a global full-precision model and allows the tiny model updates to be continuously accumulated, instead of being erased by the quantization. Furthermore, it introduces pipelining technology to parallel CPU-based quantization and DSP-enabled training, which reduces the floating-point computation overhead of frequent data quantization. Extensive experiments show that Q-FedUpdate can effectively reduce the on-device energy consumption by 21×, and accelerate the FL convergence by 6.1× with only 2% accuracy loss.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- FwdLLM: Efficient Federated Finetuning of Large Language Models with Perturbed InferencesMengwei Xu, Dongqi Cai, Yaozong Wu, Xiang Li 等USENIX ATC 2024 · 被引用 78 次
- FedMobile: Enabling Knowledge Contribution-aware Multi-modal Federated Learning with Incomplete ModalitiesYi Liu, Cong Wang, Xingliang YuanWWW 2025 · 被引用 9 次
它引用的顶会 Paper15
- Federated Learning with Matched AveragingHongyi Wang, Mikhail Yurochkin, Yuekai Sun, Dimitris S. Papailiopoulos 等ICLR 2020 · 被引用 1,368 次
- To Talk or to Work: Flexible Communication Compression for Energy Efficient Federated Learning over Heterogeneous Mobile Edge DevicesLiang Li, Dian Shi, Ronghui Hou, Hui Li 等INFOCOM 2021 · 被引用 196 次
- PyramidFL: a fine-grained client selection framework for efficient federated learningChenning Li, Xiao Zeng, Mi Zhang, Zhichao CaoMobiCom 2022 · 被引用 190 次
- Characterizing Impacts of Heterogeneity in Federated Learning upon Large-Scale Smartphone DataChengxu Yang, Qipeng Wang, Mengwei Xu, Zhenpeng Chen 等WWW 2021 · 被引用 171 次
- StageNet: Stage-Aware Neural Networks for Health Risk PredictionJunyi Gao, Cao Xiao, Yasha Wang, Wen Tang 等WWW 2020 · 被引用 131 次
相关 Paper
- Low Precision Local Training is Enough for Federated LearningZhiwei Li, Yiqiu LI, Binbin Lin, Zhongming Jin 等NeurIPS 2024 · 被引用 7 次
- ZeroFL: Efficient On-Device Training for Federated Learning with Local SparsityXinchi Qiu, Javier Fernández-Marqués, Pedro P. B. de Gusmao, Yan Gao 等ICLR 2022 · 被引用 87 次
- LBI-FL: Low-Bit Integerized Federated Learning with Temporally Dynamic Bit-Width AllocationLi Ding, Hao Zhang, Wenrui Dai, Chenglin Li 等ICML 2025
- On-Device Training Under 256KB MemoryJi Lin, Ligeng Zhu, Wei-Ming Chen, Wei-Chen Wang 等NeurIPS 2022 · 被引用 345 次
- QSFL: A Two-Level Uplink Communication Optimization Framework for Federated LearningLiping Yi, Gang Wang, Xiaoguang LiuICML 2022 · 被引用 34 次
