Fast Training Method for Stochastic Compositional Optimization Problems
Hongchang Gao, Heng Huang
摘要
The stochastic compositional optimization problem covers a wide range of machine learning models, such as sparse additive models and model-agnostic meta-learning. Thus, it is necessary to develop efficient methods for its optimization. Existing methods for the stochastic compositional optimization problem only focus on the single machine scenario, which is far from satisfactory when data are distributed on different devices. To address this problem, we propose novel decentralized stochastic compositional gradient descent methods to efficiently train the largescale stochastic compositional optimization problem. To the best of our knowledge, our work is the first one facilitating decentralized training for this kind of problem. Furthermore, we provide the convergence analysis for our methods, which shows that the convergence rate of our methods can achieve linear speedup with respect to the number of devices. At last, we apply our decentralized training methods to the model-agnostic meta-learning problem, and the experimental results confirm the superior performance of our methods.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper9
- On the Convergence of Local Stochastic Compositional Gradient Descent with MomentumHongchang Gao, Junyi Li, Heng HuangICML 2022 · 被引用 18 次
- Federated Compositional Deep AUC MaximizationXinwen Zhang, Yihan Zhang, Tianbao Yang, Richard Souvenir 等NeurIPS 2023 · 被引用 17 次
- Efficient Decentralized Stochastic Gradient Descent Method for Nonconvex Finite-Sum Optimization ProblemsWenkang Zhan, Gang Wu, Hongchang GaoAAAI 2022 · 被引用 8 次
- On The Surprising Effectiveness of a Single Global Merging in Decentralized LearningTongtian Zhu, Tianyu Zhang, Mingze Wang, Zhanpeng Zhou 等ICLR 2026 · 被引用 2 次
- On the Convergence of Stochastic Smoothed Multi-Level Compositional Gradient Descent AscentXinwen Zhang, Hongchang GaoNeurIPS 2025 · 被引用 1 次
它引用的顶会 Paper3
- Decentralized Deep Learning with Arbitrary Communication CompressionAnastasia Koloskova, Tao Lin, Sebastian U. Stich, Martin JaggiICLR 2020 · 被引用 263 次
- On the Convergence of Communication-Efficient Local SGD for Federated LearningHongchang Gao, An Xu, Heng HuangAAAI 2021 · 被引用 66 次
- Improving the Sample and Communication Complexity for Decentralized Non-Convex Optimization: Joint Gradient Estimation and TrackingHaoran Sun, Songtao Lu, Mingyi HongICML 2020 · 被引用 57 次
相关 Paper
- Stability and Generalization of Stochastic Compositional Gradient Descent AlgorithmsMing Yang, Xiyuan Wei, Tianbao Yang, Yiming YingICML 2024 · 被引用 4 次
- A Doubly Recursive Stochastic Compositional Gradient Descent Method for Federated Multi-Level Compositional OptimizationHongchang GaoICML 2024 · 被引用 1 次
- A Stochastic Linearized Augmented Lagrangian Method for Decentralized Bilevel OptimizationSongtao Lu, Siliang Zeng, Xiaodong Cui, Mark S. Squillante 等NeurIPS 2022 · 被引用 29 次
- Distributed Stochastic -Level Optimization Over NetworksXinwen Zhang, Yihan Zhang, Hongchang Gao, Heng HuangICML 2026
- Biased Stochastic First-Order Methods for Conditional Stochastic Optimization and Applications in Meta LearningYifan Hu, Siqi Zhang, Xin Chen, Niao HeNeurIPS 2020 · 被引用 69 次
