The Cost of Learning Under Multiple Change Points
Tomer Gafni, Garud Iyengar, Assaf Zeevi
摘要
We consider an online learning problem in environments with multiple change points. In contrast to the single change point problem that is widely studied using classical "high confidence" detection schemes, the multiple change point environment presents new learning-theoretic and algorithmic challenges. Specifically, we show that classical methods may exhibit catastrophic failure (high regret) due to a phenomenon we refer to as endogenous confounding. To overcome this, we propose a new class of learning algorithms dubbed Anytime Tracking CUSUM (ATC). These are horizon-free online algorithms that implement a selective detection principle, balancing the need to ignore "small" (hard-to-detect) shifts, while reacting "quickly" to significant ones. We prove that the performance of a properly tuned ATC algorithm is nearly minimax-optimal; its regret is guaranteed to closely match a novel information-theoretic lower bound on the achievable performance of any learning algorithm in the multiple change point problem. Experiments on synthetic as well as real-world data validate the aforementioned theoretical findings.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper3
- An Adaptive Deep RL Method for Non-Stationary Environments with Piecewise Stable ContextXiaoyu Chen, Xiangming Zhu, Yufeng Zheng, Pushi Zhang 等NeurIPS 2022 · 被引用 24 次
- Efficient Non-stationary Online Learning by Wavelets with Applications to Online Distribution Shift AdaptationYu-Yang Qian, Peng Zhao, Yu-Jie Zhang, Masashi Sugiyama 等ICML 2024 · 被引用 10 次
- Almost Minimax Optimal Best Arm Identification in Piecewise Stationary Linear BanditsYunlong Hou, Vincent Y. F. Tan, Zixin ZhongNeurIPS 2024 · 被引用 6 次
相关 Paper
- Change Point Detection via Multivariate Singular Spectrum AnalysisArwa Alanqary, Abdullah Omar Alomar, Devavrat ShahNeurIPS 2021 · 被引用 23 次
- Detecting and Adapting to Irregular Distribution Shifts in Bayesian Online LearningAodong Li, Alex Boyd, Padhraic Smyth, Stephan MandtNeurIPS 2021 · 被引用 31 次
- Online Label Shift: Optimal Dynamic Regret meets Practical AlgorithmsDheeraj Baby, Saurabh Garg, Tzu-Ching Yen, Sivaraman Balakrishnan 等NeurIPS 2023 · 被引用 17 次
- Adapting to Online Label Shift with Provable GuaranteesYong Bai, Yu-Jie Zhang, Peng Zhao, Masashi Sugiyama 等NeurIPS 2022 · 被引用 43 次
- Non-Stationary Lipschitz BanditsNicolas Nguyen, Solenne Gaucher, Claire VernadeNeurIPS 2025 · 被引用 3 次
