FLM-TopK: Expediting Federated Large Language Model Tuning by Sparsifying Intervalized Gradients
Wenqi Qiu, Yipeng Zhou, Jinzhi Wang, Quan Z. Sheng, Laizhong Cui
摘要
The past few years have witnessed the unprecedented capability of large language models (LLMs). To adapt LLMs with various downstream tasks, fine-tuning methods, e.g., Low-Rank Adaptation (LoRA), are proposed to efficiently tune LLMs. Meanwhile, federated LLM tuning emerges for refining LLMs with clients owning private data. In the federated tuning process, the server and clients frequently exchange fine-tune gradients via Internet, giving the rise of the communication challenge. To overcome this challenge, most existing works employ quantization methods for compressing gradients because sparsification methods like TopK incur heavy overhead for transmitting position IDs (PIDs) of sparsified gradients. In this work, to expedite federated LLM tuning with a higher compression rate, we design the Federated LLM Tuning with TopK (FLM-TopK) algorithm. Specifically, FLM-TopK intervalizes gradients before compression. Then, TopK is separately applied for gradients in each interval so that the overhead representing PIDs is constrained. To optimize our algorithm, we empirically study the distribution of gradients, which obeys the Gaussian distribution. Based on the Gaussian distribution, we establish an optimization problem to minimize the compression error by jointly optimizing the interval size and the sparsification rate per interval. We prove that the non-convex problem can be approximately solved by alternating optimization. To demonstrate the superiority of FLM-TopK, we conduct extensive experiments on nine public datasets. The results demonstrate that FLM-TopK significantly outperforms SOTA baselines, achieving 6.42%-18.87% improvement in accuracy and 17.07%-44.44% reduction in communication traffic.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
引用它的顶会 Paper4
- PrivTune: Efficient and Privacy-Preserving Fine-Tuning of Large Language Models via Device-Cloud CollaborationYi Liu, Weixiang Han, Chengjun Cai, Xingliang Yuan 等INFOCOM 2026 · 被引用 2 次
- Less Is More in Federated Continual Learning: RieSelect for Conflict-Aware Layer Selection in LLMsWenqi Qiu, Yipeng Zhou, Lin Zhu, Laizhong CuiICML 2026
- ZorBA: Zeroth-order Federated Fine-tuning of LLMs with Heterogeneous Block ActivationChuiyang Meng, Ming Tang, Vincent W. S. WongINFOCOM 2026
- Dual-Phase Federated Deep Unlearning via Weight-Aware Rollback and ReconstructionChangjun Zhou, Jintao Zheng, Leyou Yang, Pengfei WangINFOCOM 2026
相关 Paper
- FedSRD: Sparsify-Reconstruct-Decompose for Communication-Efficient Federated Large Language Models Fine-TuningGuochen Yan, Luyuan Xie, Qingni Shen, Yuejian Fang 等WWW 2026 · 被引用 1 次
- EcoLoRA: Communication-Efficient Federated Fine-Tuning of Large Language ModelsHan Liu, Ruoyao Wen, Srijith Nair, Jia Liu 等EMNLP 2025 · 被引用 2 次
- FLoRG: Federated Fine-tuning with Low-rank Gram Matrices and Procrustes AlignmentChuiyang Meng, Ming Tang, Vincent W. S. WongICLR 2026 · 被引用 7 次
- Towards Robust and Efficient Federated Low-Rank Adaptation with Heterogeneous ClientsJabin Koo, Minwoo Jang, Jungseul OkACL 2025
- Federated Sketching LoRA: A Flexible Framework for Heterogeneous Collaborative Fine-Tuning of LLMsWenzhi Fang, Dong-Jun Han, Liangqi Yuan, Seyyedali Hosseinalipour 等ICML 2026 · 被引用 4 次
