Fine-Grained Theoretical Analysis of Federated Zeroth-Order Optimization
Jun Chen, Hong Chen, Bin Gu, Hao Deng
Abstract
Federated zeroth-order optimization (FedZO) algorithm enjoys the advantages of both zeroth-order optimization and federated learning, and has shown exceptional performance on black-box attack and softmax regression tasks. However, there is little generalization analysis for FedZO, and its analysis on computing convergence rate is slower than the corresponding first-order optimization setting. This paper aims to establish systematic theoretical assessments of FedZO by developing the analysis technique of on-average model stability. We establish the first generalization error bound of FedZO under the Lipschitz continuity and smoothness conditions. Then, refined generalization and optimization bounds are provided by replacing bounded gradient with heavy-tailed gradient noise and utilizing the second-order Taylor expansion for gradient approximation. With the help of a new error decomposition strategy, our theoretical analysis is also extended to the asynchronous case. For FedZO, our fine-grained analysis fills the theoretical gap on the generalization guarantees and polishes the convergence characterization of the computing algorithm.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers8
- Federated Full-Parameter Tuning of Billion-Sized Language Models with Communication Cost under 18 KilobytesZhen Qin, Daoyuan Chen, Bingchen Qian, Bolin Ding et al.ICML 2024 · 73 citations
- Stability and Generalization of Zeroth-Order Decentralized Stochastic Gradient Descent with Changing TopologyXiaolin Hu, Zixuan Gong, Gengze Xu, Wei Liu et al.AAAI 2025 · 3 citations
- Mitigating Non-IID Drift in Zeroth-Order Federated LLM Fine-Tuning with Transferable SparsityYide Ran, Wentao Guo, Jingwei Sun, Yanzhou Pan et al.ICLR 2026 · 1 citation
- How Does Black-Box Impact the Learning Guarantee of Stochastic Compositional Optimization?Jun Chen, Hong Chen, Bin GuNeurIPS 2024 · 1 citation
- Stability and Generalization Analysis of Decentralized SGD: Sharper Bounds Beyond Lipschitzness and SmoothnessShuang Zeng, Yunwen LeiICML 2025
Builds on9
- The Distributed Discrete Gaussian Mechanism for Federated Learning with Secure AggregationPeter Kairouz, Ziyu Liu, Thomas SteinkeICML 2021 · 291 citations
- Fine-Grained Analysis of Stability and Generalization for Stochastic Gradient DescentYunwen Lei, Yiming YingICML 2020 · 165 citations
- Sharper Generalization Bounds for Learning with Gradient-dominated Objective FunctionsYunwen Lei, Yiming YingICLR 2021 · 52 citations
- Sharper Generalization Bounds for Pairwise LearningYunwen Lei, Antoine Ledent, Marius KloftNeurIPS 2020 · 50 citations
- High Probability Guarantees for Nonconvex Stochastic Gradient Descent with Heavy TailsShaojie Li, Yong LiuICML 2022 · 37 citations
Related papers
- Towards Understanding Generalization of Federated Adversarial Learning: Perspective of Algorithmic StabilityYongkang Yang, Chang Cao, Ke Zhang, Han Li et al.ICML 2026
- Black-Box Generalization: Stability of Zeroth-Order LearningKonstantinos E. Nikolakakis, Farzin Haddadpour, Dionysios S. Kalogerias, Amin KarbasiNeurIPS 2022
- Sharper Generalization Guarantees for Asynchronous SGD: Beyond Lipschitzness, Smoothness and Data HomogeneityYufeng Xie, Yunwen LeiICML 2026
- Tight High-Probability Bounds for Nonconvex Heavy-Tailed Scenario under Weaker AssumptionsWeixin An, Yuanyuan Liu, Fanhua Shang, Han Yu et al.NeurIPS 2025
- General Stability Analysis for Zeroth-Order Optimization AlgorithmsXinyue Liu, Hualin Zhang, Bin Gu, Hong ChenICLR 2024 · 3 citations
