Bitwidth Heterogeneous Federated Learning with Progressive Weight Dequantization
Jaehong Yoon, Geon Park, Wonyong Jeong, Sung Ju Hwang
摘要
In practical federated learning scenarios, the participating devices may have different bitwidths for computation and memory storage by design. However, despite the progress made in device-heterogeneous federated learning scenarios, the heterogeneity in the bitwidth specifications in the hardware has been mostly overlooked. We introduce a pragmatic FL scenario with bitwidth heterogeneity across the participating devices, dubbed as Bitwidth Heterogeneous Federated Learning (BHFL). BHFL brings in a new challenge, that the aggregation of model parameters with different bitwidths could result in severe performance degeneration, especially for high-bitwidth models. To tackle this problem, we propose ProWD framework, which has a trainable weight dequantizer at the central server that progressively reconstructs the low-bitwidth weights into higher bitwidth weights, and finally into full-precision weights. ProWD further selectively aggregates the model parameters to maximize the compatibility across bit-heterogeneous weights. We validate ProWD against relevant FL baselines on the benchmark datasets, using clients with varying bitwidths. Our ProWD largely outperforms the baseline FL algorithms as well as naive approaches (e.g. grouped averaging) under the proposed BHFL scenario.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper7
- Conformal Prediction for Federated Uncertainty Quantification Under Label ShiftVincent Plassier, Mehdi Makni, Aleksandr Rubashevskii, Eric Moulines 等ICML 2023 · 被引用 28 次
- Recurrent Early Exits for Federated Learning with Heterogeneous ClientsRoyson Lee, Javier Fernández-Marqués, Shell Xu Hu, Da Li 等ICML 2024 · 被引用 13 次
- Fed-QSSL: A Framework for Personalized Federated Learning under Bitwidth and Data HeterogeneityYiyue Chen, Haris Vikalo, Chianing WangAAAI 2024 · 被引用 13 次
- Towards Energy-efficient Federated Learning via INT8-based Training on Mobile DSPsJinliang Yuan, Shangguang Wang, Hongyu Li, Daliang Xu 等WWW 2024 · 被引用 8 次
- Caesar: Optimizing Federated Learning via Low-deviation CompressionJiaming Yan, Jianchun Liu, Hongli Xu, Zhenguo Ma 等KDD 2026 · 被引用 1 次
它引用的顶会 Paper11
- Ensemble Distillation for Robust Model Fusion in Federated LearningTao Lin, Lingjing Kong, Sebastian U. Stich, Martin JaggiNeurIPS 2020 · 被引用 1,615 次
- Federated Learning with Matched AveragingHongyi Wang, Mikhail Yurochkin, Yuekai Sun, Dimitris S. Papailiopoulos 等ICLR 2020 · 被引用 1,368 次
- Group Knowledge Transfer: Federated Learning of Large CNNs at the EdgeChaoyang He, Murali Annavaram, Salman AvestimehrNeurIPS 2020 · 被引用 605 次
- Federated Continual Learning with Weighted Inter-client TransferJaehong Yoon, Wonyong Jeong, Giwoong Lee, Eunho Yang 等ICML 2021 · 被引用 303 次
- Ultra-Low Precision 4-bit Training of Deep Neural NetworksXiao Sun, Naigang Wang, Chia-Yu Chen, Jiamin Ni 等NeurIPS 2020 · 被引用 227 次
相关 Paper
- Mixed-Precision Quantization for Federated Learning on Resource-Constrained Heterogeneous DevicesHuancheng Chen, Haris VikaloCVPR 2024
- DynFed: Adaptive Federated Learning via Quantization-Aware Knowledge DistillationNan He, Yiming Chen, Zheng Jiang, Song Yang 等ACM MM 2025 · 被引用 1 次
- FedWSQ: Efficient Federated Learning with Weight Standardization and Distribution-Aware Non-Uniform QuantizationSeung-Wook Kim, Seongyeol Kim, Jiah Kim, Seowon Ji 等ICCV 2025 · 被引用 3 次
- HADFL: Heterogeneity-aware Decentralized Federated Learning FrameworkJing Cao, Zirui Lian, Weihong Liu, Zongwei Zhu 等DAC 2021 · 被引用 28 次
- Low Precision Local Training is Enough for Federated LearningZhiwei Li, Yiqiu LI, Binbin Lin, Zhongming Jin 等NeurIPS 2024 · 被引用 7 次
