Lune

NeurIPS2023顶会

On Robust Streaming for Learning with Experts: Algorithms and Lower Bounds

David P. Woodruff, Fred Zhang, Samson Zhou

2023年份
7被引次数
5顶会引用

摘要

In the online learning with experts problem, an algorithm makes predictions about an outcome on each of T days, given a set of n experts who make predictions on each day. The algorithm is given feedback on the outcomes of each day, including the cost of its prediction and the cost of the expert predictions, and the goal is to make a prediction with the minimum cost, compared to the best expert in hindsight. However, often the predictions made by experts or algorithms at some time influence future outcomes, so that the input is adaptively generated. In this paper, we study robust algorithms for the experts problem under memory constraints. We first give a randomized algorithm that is robust to adaptive inputs that uses (cid:101) O (cid:16) nR √ T (cid:17) space for regret R when the best expert makes M = O (cid:16) R 2 T log 2 n (cid:17) mistakes, thereby showing a smooth space-regret trade-off. We then show a space lower bound of (cid:101) Ω (cid:0) nMRT (cid:1) for any randomized algorithm that achieves regret R with probability 1 − 2 − Ω( T ) . Such an algorithm is useful for adaptive inputs, as the failure probability is low enough to union bound over all computation paths. Our result implies that the natural deterministic algorithm, which iterates through pools of experts until each expert in the pool has erred, is optimal up to polylogarithmic factors. Finally, we empirically demonstrate the benefit of using robust procedures against a white-box adversary that has access to the internal state of the algorithm.

问问这篇 Paper

智能体会读完全文。

Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。

可以从这些问题问起

智能体调用

Luneget_paper_fulltext

在 Lune 里问

免费开始,无需绑卡

引用它的顶会 Paper5

问问它们各自怎么用它

它引用的顶会 Paper15

相关 Paper

黄昏的海面,两侧是细线勾勒的悬崖