Generalized Implicit Follow-The-Regularized-Leader
Keyi Chen, Francesco Orabona
Abstract
We propose a new class of online learning algorithms, generalized implicit Follow-The-Regularized-Leader (FTRL), that expands the scope of FTRL framework. Generalized implicit FTRL can recover known algorithms, as FTRL with linearized losses and implicit FTRL, and it allows the design of new update rules, as extensions of aProx and Mirror-Prox to FTRL. Our theory is constructive in the sense that it provides a simple unifying framework to design updates that directly improve the worst-case upper bound on the regret. The key idea is substituting the linearization of the losses with a Fenchel-Young inequality. We show the flexibility of the framework by proving that some known algorithms, like the Mirror-Prox updates, are instantiations of the generalized implicit FTRL. Finally, the new framework allows us to recover the temporal variation bound of implicit OMD, with the same computational complexity.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext cc0d5b13-12e3-4988-a49c-c9d63b670356Cited by top-tier papers2
- Optimistic Online Mirror Descent for Bridging Stochastic and Adversarial Online Convex OptimizationSijia Chen, Wei-Wei Tu, Peng Zhao, Lijun ZhangICML 2023 · 33 citations
- Implicit Riemannian Optimism with Applications to Min-Max ProblemsChristophe Roux, David Martínez-Rubio, Sebastian PokuttaICML 2025
Builds on3
- Online mirror descent and dual averaging: keeping pace in the dynamic caseHuang Fang, Nick Harvey, Victor S. Portella, Michael P. FriedlanderICML 2020 · 38 citations
- Temporal Variability in Implicit Online LearningNicolò Campolongo, Francesco OrabonaNeurIPS 2020 · 29 citations
- Better Parameter-Free Stochastic Optimization with ODE Updates for Coin-BettingKeyi Chen, John Langford, Francesco OrabonaAAAI 2022 · 23 citations
Related papers
- Equivalence Analysis between Counterfactual Regret Minimization and Online Mirror DescentWeiming Liu, Huacong Jiang, Bin Li, Houqiang LiICML 2022 · 13 citations
- (Doubly) Exponential Lower Bounds for Follow the Regularized Leader in Potential GamesIoannis Anagnostides, Ioannis Panageas, Nikolas Patris, Tuomas SandholmICML 2026 · 2 citations
- On the Dynamic Regret of Following the Regularized Leader: Optimism with History PruningNaram Mhaisen, George IosifidisICML 2025
- Best-case lower bounds in online learningCristóbal Guzmán, Nishant A. Mehta, Ali MortazaviNeurIPS 2021 · 2 citations
- Faster Game Solving via Predictive Blackwell Approachability: Connecting Regret Matching and Mirror DescentGabriele Farina, Christian Kroer, Tuomas SandholmAAAI 2021 · 91 citations
