Stabilizing Linear Passive-Aggressive Online Learning with Weighted Reservoir Sampling
Skyler Wu, Fred Lu, Edward Raff, James Holt
摘要
Online learning methods, like the seminal Passive-Aggressive (PA) classifier, are still highly effective for high-dimensional streaming data, out-of-core processing, and other throughput-sensitive applications. Many such algorithms rely on fast adaptation to individual errors as a key to their convergence. While such algorithms enjoy low theoretical regret, in real-world deployment they can be sensitive to individual outliers that cause the algorithm to over-correct. When such outliers occur at the end of the data stream, this can cause the final solution to have unexpectedly low accuracy. We design a weighted reservoir sampling (WRS) approach to obtain a stable ensemble model from the sequence of solutions without requiring additional passes over the data, hold-out sets, or a growing amount of memory. Our key insight is that good solutions tend to be error-free for more iterations than bad solutions, and thus, the number of passive rounds provides an estimate of a solution's relative quality. Our reservoir thus contains previous intermediate weight vectors with high survival times. We demonstrate our WRS approach on the Passive-Aggressive Classifier (PAC) and First-Order Sparse Online Learning (FSOL), where our method consistently and significantly outperforms the unmodified approach. We show that the risk of the ensemble classifier is bounded with respect to the regret of the underlying online learning method.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
相关 Paper
- Adaptive Passive-Aggressive Framework for Online Regression with Side InformationRunhao Shi, Jiaxi Ying, Daniel P. PalomarNeurIPS 2024 · 被引用 3 次
- Online Continual Learning from Imbalanced DataAristotelis Chrysakis, Marie-Francine MoensICML 2020 · 被引用 166 次
- One-Pass Diversified Sampling with Application to Terabyte-Scale Genomic Sequence StreamsBenjamin Coleman, Benito Geordie, Li Chou, Ryan A. Leo Elworth 等ICML 2022 · 被引用 11 次
- Information-theoretic Online Memory Selection for Continual LearningShengyang Sun, Daniele Calandriello, Huiyi Hu, Ang Li 等ICLR 2022 · 被引用 61 次
- ReservoirTTA: Prolonged Test-time Adaptation for Evolving and Recurring DomainsGuillaume Vray, Devavrat Tomar, Xufeng Gao, Jean-Philippe Thiran 等NeurIPS 2025 · 被引用 9 次
