Nearly Optimal Bounds for Cyclic Forgetting
William Swartworth, Deanna Needell, Rachel A. Ward, Mark Kong, Halyun Jeong
摘要
We provide theoretical bounds on the forgetting quantity in the continual learning setting for linear tasks, where each round of learning corresponds to projecting onto a linear subspace. For a cyclic task ordering on T tasks repeated m times each, we prove the best known upper bound of O ( T 2 /m ) on the forgetting. Notably, our bound holds uniformly over all choices of tasks and is independent of the ambient dimension. Our main technical contribution is a characterization of the union of all numerical ranges of products of T (real or complex) projections as a sinusoidal spiral, which may be of independent interest
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper10
- The Joint Effect of Task Similarity and Overparameterization on Catastrophic Forgetting - An Analytical ModelDaniel Goldfarb, Itay Evron, Nir Weinberger, Daniel Soudry 等ICLR 2024 · 被引用 25 次
- Understanding Forgetting in Continual Learning with Linear RegressionMeng Ding, Kaiyi Ji, Di Wang, Jinhui XuICML 2024 · 被引用 23 次
- Fast Last-Iterate Convergence of SGD in the Smooth Interpolation RegimeAmit Attia, Matan Schliserman, Uri Sherman, Tomer KorenNeurIPS 2025 · 被引用 18 次
- Optimal Rates in Continual Linear Regression via Increasing RegularizationRan Levinstein, Amit Attia, Matan Schliserman, Uri Sherman 等NeurIPS 2025 · 被引用 10 次
- Are Greedy Task Orderings Better Than Random in Continual Linear Regression?Matan Tsipory, Ran Levinstein, Itay Evron, Mark Kong 等NeurIPS 2025 · 被引用 5 次
它引用的顶会 Paper2
相关 Paper
- Continual Learning in Linear Classification on Separable DataItay Evron, Edward Moroshko, Gon Buzaglo, Maroun Khriesh 等ICML 2023 · 被引用 32 次
- Convergence and Implicit Bias of Gradient Descent on Continual Linear ClassificationHyunji Jung, Hanseul Cho, Chulhee YunICLR 2025
- Adaptive Orthogonal Projection for Batch and Online Continual LearningYiduo Guo, Wenpeng Hu, Dongyan Zhao, Bing LiuAAAI 2022 · 被引用 56 次
- Optimizing Spca-based Continual Learning: A Theoretical ApproachChunchun Yang, Malik Tiomoko, Zengfu WangICLR 2023
- On the Theory of Continual Learning with Gradient Descent for Neural NetworksHossein Taheri, Avishek Ghosh, Arya MazumdarICML 2026 · 被引用 2 次
