Lune

ICML2023顶会

Escaping saddle points in zeroth-order optimization: the power of two-point estimators

Zhaolin Ren, Yujie Tang, Na Li

2023年份
13被引次数
3顶会引用

摘要

Two-point zeroth order methods are important in many applications of zeroth-order optimization, such as robotics, wind farms, power systems, online optimization, and adversarial robustness to black-box attacks in deep neural networks, where the problem may be high-dimensional and/or time-varying. Most problems in these applications are nonconvex and contain saddle points. While existing works have shown that zeroth-order methods utilizing Ω(d)\Omega(d) function valuations per iteration (with dd denoting the problem dimension) can escape saddle points efficiently, it remains an open question if zeroth-order methods based on two-point estimators can escape saddle points. In this paper, we show that by adding an appropriate isotropic perturbation at each iteration, a zeroth-order algorithm based on 2m2m (for any 1≤m≤d1 \leq m \leq d) function evaluations per iteration can not only find ϵ\epsilon-second order stationary points polynomially fast, but do so using only O~(dmϵ2ψˉ)\tilde{O}\left(\frac{d}{m\epsilon^{2}\bar{\psi}}\right) function evaluations, where ψˉ≥Ω~(ϵ)\bar{\psi} \geq \tilde{\Omega}\left(\sqrt{\epsilon}\right) is a parameter capturing the extent to which the function of interest exhibits the strict saddle property.

问问这篇 Paper

智能体会读完全文。

Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。

可以从这些问题问起

智能体调用

Luneget_paper_fulltext

在 Lune 里问

免费开始,无需绑卡

引用它的顶会 Paper3

问问它们各自怎么用它

它引用的顶会 Paper7

相关 Paper

黄昏的海面,两侧是细线勾勒的悬崖