Lune

ICLR2026顶会

Nesterov Finds GRAAL: Optimal and Adaptive Gradient Method for Convex Optimization

Ekaterina Borodich, Dmitry Kovalev

2026年份
10被引次数
2顶会引用

摘要

In this paper, we focus on the problem of minimizing a continuously differentiable convex objective function, min⁡xf(x)\min_x f(x). Recently, Malitsky (2020); Alacaoglu et al.(2023) developed an adaptive first-order method, GRAAL. This algorithm computes stepsizes by estimating the local curvature of the objective function without any line search procedures or hyperparameter tuning, and attains the standard iteration complexity O(L∥x0−x∗∥2/ϵ)\mathcal{O}(L\lVert x_0-x^*\rVert^2/\epsilon) of fixed-stepsize gradient descent for LL-smooth functions. However, a natural question arises: is it possible to accelerate the convergence of GRAAL to match the optimal complexity O(L∥x0−x∗∥2/ϵ)\mathcal{O}(\sqrt{L\lVert x_0-x^*\rVert^2/\epsilon}) of the accelerated gradient descent of Nesterov (1983)? Although some attempts have been made by Li and Lan (2025); Suh and Ma (2025), the ability of existing accelerated algorithms to adapt to the local curvature of the objective function is highly limited. We resolve this issue and develop GRAAL with Nesterov acceleration, which can adapt its stepsize to the local curvature at a geometric, or linear, rate just like non-accelerated GRAAL. We demonstrate the adaptive capabilities of our algorithm by proving that it achieves near-optimal iteration complexities for LL-smooth functions, as well as under a more general (L0,L1)(L_0,L_1)-smoothness assumption (Zhang et al., 2019).

问问这篇 Paper

智能体会读完全文。

Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。

可以从这些问题问起

智能体调用

Luneget_paper_fulltext

在 Lune 里问

免费开始,无需绑卡

lune papers fulltext dba3f7d1-8adc-400d-94d4-c4d9745d8ffb

引用它的顶会 Paper2

问问它们各自怎么用它

它引用的顶会 Paper15

相关 Paper

黄昏的海面,两侧是细线勾勒的悬崖