Learning Rate Schedules in the Presence of Distribution Shift
Matthew Fahrbach, Adel Javanmard, Vahab Mirrokni, Pratik Worah
摘要
We design learning rate schedules that minimize regret for SGD-based online learning in the presence of a changing data distribution. We fully characterize the optimal learning rate schedule for online linear regression via a novel analysis with stochastic differential equations. For general convex loss functions, we propose new learning rate schedules that are robust to distribution shift, and we give upper and lower bounds for the regret that only differ by constants. For non-convex loss functions, we define a notion of regret based on the gradient norm of the estimated models and propose a learning schedule that minimizes an upper bound on the total expected regret. Intuitively, one expects changing loss landscapes to require more exploration, and we confirm that optimal learning rate schedules typically increase in the presence of distribution shift. Finally, we provide experiments for high-dimensional regression models and neural networks to illustrate these learning rate schedules and their cumulative regret.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- Unified Embedding: Battle-Tested Feature Representations for Web-Scale ML SystemsBenjamin Coleman, Wang-Cheng Kang, Matthew Fahrbach, Ruoxi Wang 等NeurIPS 2023 · 被引用 30 次
- Understanding the Robustness of Multi-modal Contrastive Learning to Distribution ShiftYihao Xue, Siddharth Joshi, Dang Nguyen, Baharan MirzasoleimanICLR 2024 · 被引用 6 次
- PriorBoost: An Adaptive Algorithm for Learning from Aggregate ResponsesAdel Javanmard, Matthew Fahrbach, Vahab MirrokniICML 2024 · 被引用 6 次
- An Online Adaptive Sampling Algorithm for Stochastic Difference-of-convex Optimization with Time-varying DistributionsYuhan Ye, Ying Cui, Jingyi WangICML 2025
它引用的顶会 Paper2
相关 Paper
- Online Adaptation to Label Distribution ShiftRuihan Wu, Chuan Guo, Yi Su, Kilian Q. WeinbergerNeurIPS 2021 · 被引用 77 次
- Label Shift Meets Online Learning: Ensuring Consistent Adaptation with Universal Dynamic RegretYucong Dai, Shilin Gu, Ruidong Fan, Chao Xu 等CVPR 2025
- Improved Online Conformal Prediction via Strongly Adaptive Online LearningAadyot Bhatnagar, Huan Wang, Caiming Xiong, Yu BaiICML 2023 · 被引用 87 次
- Optimal Margin Distribution Learning in Dynamic EnvironmentsTeng Zhang, Peng Zhao, Hai JinAAAI 2020 · 被引用 7 次
- Online Label Shift: Optimal Dynamic Regret meets Practical AlgorithmsDheeraj Baby, Saurabh Garg, Tzu-Ching Yen, Sivaraman Balakrishnan 等NeurIPS 2023 · 被引用 17 次
