Instance-Conditional Timescales of Decay for Non-Stationary Learning
Nishant Jain, Pradeep Shenoy
摘要
Slow concept drift is a ubiquitous, yet under-studied problem in practical machine learning systems. In such settings, although recent data is more indicative of future data, naively prioritizing recent instances runs the risk of losing valuable information from the past. We propose an optimization-driven approach towards balancing instance importance over large training windows. First, we model instance relevance using a mixture of multiple timescales of decay, allowing us to capture rich temporal trends. Second, we learn an auxiliary scorer model that recovers the appropriate mixture of timescales as a function of the instance itself. Finally, we propose a nested optimization objective for learning the scorer, by which it maximizes forward transfer for the learned model. Experiments on a large real-world dataset of 39M photos over a 9 year period show upto 15% relative gains in accuracy compared to other robust learning baselines. We replicate our gains on two collections of real-world datasets for non-stationary learning, and extend our work to continual learning settings where, too, we beat SOTA methods by large margins.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- Learning model uncertainty as variance-minimizing instance weightsNishant Jain, Karthikeyan Shanmugam, Pradeep ShenoyICLR 2024 · 被引用 7 次
- Improving Generalization via Meta-Learning on Hard SamplesNishant Jain, Arun Sai Suggala, Pradeep ShenoyCVPR 2024
它引用的顶会 Paper10
- Adversarial Domain Adaptation with Domain MixupMinghao Xu, Jian Zhang, Bingbing Ni, Teng Li 等AAAI 2020 · 被引用 499 次
- Efficient and Modular Implicit DifferentiationMathieu Blondel, Quentin Berthet, Marco Cuturi, Roy Frostig 等NeurIPS 2022 · 被引用 386 次
- Improving Out-of-Distribution Robustness via Selective AugmentationHuaxiu Yao, Yu Wang, Sai Li, Linjun Zhang 等ICML 2022 · 被引用 275 次
- Continual Prototype Evolution: Learning Online from Non-Stationary Data StreamsMatthias De Lange, Tinne TuytelaarsICCV 2021 · 被引用 251 次
- Prioritized Training on Points that are Learnable, Worth Learning, and not yet LearntSören Mindermann, Jan Markus Brauner, Muhammed Razzak, Mrinank Sharma 等ICML 2022 · 被引用 237 次
相关 Paper
- DeepBooTS: Dual-Stream Residual Boosting for Drift-Resilient Time-Series ForecastingDaojun Liang, Jing Chen, Xiao Wang, Yinglong Wang 等AAAI 2026
- TRACE: A Generalizable Drift Detector for Streaming Data-Driven OptimizationYuan-Ting Zhong, Ting Huang, Xiaolin Xiao, Yue-Jiao GongAAAI 2026 · 被引用 1 次
- Training for the Future: A Simple Gradient Interpolation Loss to Generalize Along TimeAnshul Nasery, Soumyadeep Thakur, Vihari Piratla, Abir De 等NeurIPS 2021 · 被引用 42 次
- Minimax Classification under Concept Drift with Multidimensional Adaptation and Performance GuaranteesVerónica Álvarez, Santiago Mazuelas, José Antonio LozanoICML 2022 · 被引用 6 次
- The AdEMAMix Optimizer: Better, Faster, OlderMatteo Pagliardini, Pierre Ablin, David GrangierICLR 2025
