Lune

ICML2025顶会

Exploiting Curvature in Online Convex Optimization with Delayed Feedback

Hao Qiu, Emmanuel Esposito, Mengxiao Zhang

出版方
2025年份
4顶会引用

摘要

In this work, we study the online convex optimization problem with curved losses and delayed feedback. When losses are strongly convex, existing approaches obtain regret bounds of order dmax⁡ln⁡Td_{\max} \ln T, where dmax⁡d_{\max} is the maximum delay and TT is the time horizon. However, in many cases, this guarantee can be much worse than dtot\sqrt{d_{\mathrm{tot}}} as obtained by a delayed version of online gradient descent, where dtotd_{\mathrm{tot}} is the total delay. We bridge this gap by proposing a variant of follow-the-regularized-leader that obtains regret of order min⁡{σmax⁡ln⁡T,dtot}\min\{\sigma_{\max}\ln T, \sqrt{d_{\mathrm{tot}}}\}, where σmax⁡\sigma_{\max} is the maximum number of missing observations. We then consider exp-concave losses and extend the Online Newton Step algorithm to handle delays with an adaptive learning rate tuning, achieving regret min⁡{dmax⁡nln⁡T,dtot}\min\{d_{\max} n\ln T, \sqrt{d_{\mathrm{tot}}}\} where nn is the dimension. To our knowledge, this is the first algorithm to achieve such a regret bound for exp-concave losses. We further consider the problem of unconstrained online linear regression and achieve a similar guarantee by designing a variant of the Vovk-Azoury-Warmuth forecaster with a clipping trick. Finally, we implement our algorithms and conduct experiments under various types of delay and losses, showing an improved performance over existing methods.

问问这篇 Paper

智能体会读完全文。

Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。

可以从这些问题问起

智能体调用

Luneget_paper_fulltext

在 Lune 里问

免费开始,无需绑卡

引用它的顶会 Paper4

问问它们各自怎么用它

它引用的顶会 Paper14

相关 Paper

黄昏的海面,两侧是细线勾勒的悬崖