Non-Exponentially Weighted Aggregation: Regret Bounds for Unbounded Loss Functions
Pierre Alquier
摘要
We tackle the problem of online optimization with a general, possibly unbounded, loss function. It is well known that the exponentially weighted aggregation strategy (EWA) leads to a regret in after steps, under the assumption that the loss is bounded. The online gradient algorithm (OGA) has a regret in when the loss is convex and Lipschitz. In this paper, we study a generalized aggregation strategy, where the weights do no longer necessarily depend exponentially on the losses. Our strategy can be interpreted as the minimization of the expected losses plus a penalty term. When the penalty term is the Kullback-Leibler divergence, we obtain EWA as a special case, but using alternative divergences lead to a regret bounds for unbounded, not necessarily convex losses. However, the cost is a worst regret bound in some cases.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper7
- A Rigorous Link between Deep Ensembles and (Variational) Bayesian MethodsVeit David Wild, Sahra Ghalebikesabi, Dino Sejdinovic, Jeremias KnoblauchNeurIPS 2023 · 被引用 40 次
- Risk Monotonicity in Statistical LearningZakaria MhammediNeurIPS 2021 · 被引用 10 次
- Optimal Comparator Adaptive Online Learning with Switching CostZhiyu Zhang, Ashok Cutkosky, Yannis PaschalidisNeurIPS 2022 · 被引用 10 次
- Minimax Optimal Quantile and Semi-Adversarial Regret via Root-Logarithmic RegularizersJeffrey Negrea, Blair L. Bilodeau, Nicolò Campolongo, Francesco Orabona 等NeurIPS 2021 · 被引用 9 次
- Practical and Matching Gradient Variance Bounds for Black-Box Variational Bayesian InferenceKyurae Kim, Kaiwen Wu, Jisu Oh, Jacob R. GardnerICML 2023 · 被引用 8 次
它引用的顶会 Paper3
- Optimal Bounds between f-Divergences and Integral Probability MetricsRohit Agrawal, Thibaut HorelICML 2020 · 被引用 50 次
- Convergence Rates of Variational Inference in Sparse Deep LearningBadr-Eddine Chérief-AbdellatifICML 2020 · 被引用 43 次
- Provable Smoothness Guarantees for Black-Box Variational InferenceJustin DomkeICML 2020 · 被引用 41 次
相关 Paper
- Exploiting Curvature in Online Convex Optimization with Delayed FeedbackHao Qiu, Emmanuel Esposito, Mengxiao ZhangICML 2025
- A Simple yet Universal Strategy for Online Convex OptimizationLijun Zhang, Guanghui Wang, Jinfeng Yi, Tianbao YangICML 2022
- Fully Unconstrained Online LearningAshok Cutkosky, Zakaria MhammediNeurIPS 2024 · 被引用 13 次
- Unconstrained Online Learning with Unbounded LossesAndrew Jacobsen, Ashok CutkoskyICML 2023 · 被引用 25 次
- A Regret-Variance Trade-Off in Online LearningDirk van der Hoeven, Nikita Zhivotovskiy, Nicolò Cesa-BianchiNeurIPS 2022 · 被引用 9 次
