Algorithmic Instabilities of Accelerated Gradient Descent
Amit Attia, Tomer Koren
摘要
We study the algorithmic stability of Nesterov's accelerated gradient method. For convex quadratic objectives, Chen et al. ( 2018 ) proved that the uniform stability of the method grows quadratically with the number of optimization steps, and conjectured that the same is true for the general convex and smooth case. We disprove this conjecture and show, for two notions of algorithmic stability (including uniform stability), that the stability of Nesterov's accelerated method in fact deteriorates exponentially fast with the number of gradient steps. This stands in sharp contrast to the bounds in the quadratic case, but also to known results for non-accelerated gradient methods where stability typically grows linearly with the number of steps.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- High Probability Bounds for Non-Convex Stochastic Optimization with MomentumShaojie Li, Pengwei Tang, Bowei Zhu, Yong LiuICLR 2026 · 被引用 100 次
- Reproducibility in Optimization: Theoretical Framework and LimitsKwangjun Ahn, Prateek Jain, Ziwei Ji, Satyen Kale 等NeurIPS 2022 · 被引用 32 次
- Optimal Guarantees for Algorithmic Reproducibility and Gradient Complexity in Convex OptimizationLiang Zhang, Junchi Yang, Amin Karbasi, Niao HeNeurIPS 2023 · 被引用 4 次
- Accelerated Dual Method for Distributed Optimization: An Inexact-Gradient View of Local UpdatesJunchi Yang, Ziyang Zeng, Linxuan Pan, Murat Yildirim 等ICML 2026
它引用的顶会 Paper3
- Stability of Stochastic Gradient Descent on Nonsmooth Convex LossesRaef Bassily, Vitaly Feldman, Cristóbal Guzmán, Kunal TalwarNeurIPS 2020 · 被引用 240 次
- Stochastic Optimization with Laggard Data PipelinesNaman Agarwal, Rohan Anil, Tomer Koren, Kunal Talwar 等NeurIPS 2020 · 被引用 14 次
- Private stochastic convex optimization: optimal rates in linear timeVitaly Feldman, Tomer Koren, Kunal TalwarSTOC 2020 · 被引用 8 次
相关 Paper
- Acceleration via Fractal Learning Rate SchedulesNaman Agarwal, Surbhi Goel, Cyril ZhangICML 2021 · 被引用 19 次
- Convex and Non-convex Optimization Under Generalized SmoothnessHaochuan Li, Jian Qian, Yi Tian, Alexander Rakhlin 等NeurIPS 2023 · 被引用 93 次
- Stability and Sharper Risk Bounds with Convergence Rate Õ(1/n2)Bowei Zhu, Shaojie Li, Mingyang Yi, Yong LiuNeurIPS 2025 · 被引用 2 次
- Fine-Grained Analysis of Stability and Generalization for Stochastic Gradient DescentYunwen Lei, Yiming YingICML 2020 · 被引用 165 次
- On the Convergence of Nesterov's Accelerated Gradient Method in Stochastic SettingsMahmoud Assran, Mike RabbatICML 2020 · 被引用 71 次
