Biased Stochastic First-Order Methods for Conditional Stochastic Optimization and Applications in Meta Learning
Yifan Hu, Siqi Zhang, Xin Chen, Niao He
摘要
Conditional stochastic optimization covers a variety of applications ranging from invariant learning and causal inference to meta-learning. However, constructing unbiased gradient estimators for such problems is challenging due to the composition structure. As an alternative, we propose a biased stochastic gradient descent (BSGD) algorithm and study the bias-variance tradeoff under different structural assumptions. We establish the sample complexities of BSGD for strongly convex, convex, and weakly convex objectives under smooth and non-smooth conditions. Our lower bound analysis shows that the sample complexities of BSGD cannot be improved for general convex objectives and nonconvex objectives except for smooth nonconvex objectives with Lipschitz continuous gradient estimator. For this special setting, we propose an accelerated algorithm called biased SpiderBoost (BSpiderBoost) that matches the lower bound complexity. We further conduct numerical experiments on invariant logistic regression and model-agnostic meta-learning to illustrate the performance of BSGD and BSpiderBoost.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper32
- Descending through a Crowded Valley - Benchmarking Deep Learning OptimizersRobin M. Schmidt, Frank Schneider, Philipp HennigICML 2021 · 被引用 195 次
- On the Bias-Variance-Cost Tradeoff of Stochastic OptimizationYifan Hu, Xin Chen, Niao HeNeurIPS 2021 · 被引用 39 次
- Finite-Sum Coupled Compositional Stochastic Optimization: Theory and ApplicationsBokun Wang, Tianbao YangICML 2022 · 被引用 38 次
- Data-Driven Conditional Robust OptimizationAbhilash Reddy Chenreddy, Nymisha Bandi, Erick DelageNeurIPS 2022 · 被引用 34 次
- On the Convergence Theory of Debiased Model-Agnostic Meta-Reinforcement LearningAlireza Fallah, Kristian Georgiev, Aryan Mokhtari, Asuman E. OzdaglarNeurIPS 2021 · 被引用 31 次
相关 Paper
- A Guide Through the Zoo of Biased SGDYury Demidovich, Grigory Malinovsky, Igor Sokolov, Peter RichtárikNeurIPS 2023 · 被引用 56 次
- Debiasing Conditional Stochastic OptimizationLie He, Shiva Prasad KasiviswanathanNeurIPS 2023 · 被引用 7 次
- Federated Conditional Stochastic OptimizationXidong Wu, Jianhui Sun, Zhengmian Hu, Junyi Li 等NeurIPS 2023 · 被引用 5 次
- Generalized-Smooth Nonconvex Optimization is As Efficient As Smooth Nonconvex OptimizationZiyi Chen, Yi Zhou, Yingbin Liang, Zhaosong LuICML 2023 · 被引用 58 次
- Fast Training Method for Stochastic Compositional Optimization ProblemsHongchang Gao, Heng HuangNeurIPS 2021 · 被引用 17 次
