Lune

ICLR2023顶会

Stochastic No-regret Learning for General Games with Variance Reduction

Yichi Zhou, Fang Kong, Shuai Li

出版方
2023年份
1顶会引用

摘要

We show that a stochastic version of optimistic mirror descent (OMD), a variant of mirror descent with recency bias, converges fast in general games. More specifically, with our algorithm, the individual regret of each player vanishes at a speed of O(1/T3/4)O(1/T^{3/4}) and the sum of all players' regret vanishes at a speed of O(1/T)O(1/T), which is an improvement upon the O(1/T)O(1/\sqrt{T}) convergence rate of prior stochastic algorithms, where TT is the number of interaction rounds. Due to the advantage of stochastic methods in the computational cost, we significantly improve the time complexity over the deterministic algorithms to approximate coarse correlated equilibrium. To achieve lower time complexity, we equip the stochastic version of OMD in with a novel low-variance Monte-Carlo estimator. Our algorithm extends previous works from two-player zero-sum games to general games.

问问这篇 Paper

问问你的智能体。

Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。

可以从这些问题问起

智能体调用

Lunesearch_papers

在 Lune 里问

免费开始,无需绑卡

引用它的顶会 Paper1

问问它们各自怎么用它

相关 Paper

黄昏的海面,两侧是细线勾勒的悬崖