Enabling Real-Time Inference in Online Continual Learning via Device-Cloud Collaboration
Haibo Liu, Chen Gong, Zhenzhe Zheng, Shengzhong Liu, Fan Wu
摘要
Online continual learning (CL) is becoming a mainstream paradigm to learn incrementally from task streams without forgetting previously learned knowledge. However, the current online CL primarily focuses on learning performance, such as avoiding catastrophic forgetting, neglecting the critical demands of system performance, such as real-time inference. As a result, the performance of real-time inference in online CL degrades significantly due to frequent data distribution variations and time-consuming model adaptation. In this work, we propose ELITE, an online CL framework with device-cloud collaboration, to realize on-device real-time inference on time-varying task streams with performance guarantee. To realize on-device real-time inference in online CL, ELITE features a new design of the model zoo comprising various pre-trained models with the assistance of the cloud, and proposes a task-oriented on-device model selection to quickly retrieve the best-fit models instead of performing time-consuming model retraining. To prevent performance degradation on new tasks not available in the cloud, we introduces a latency-aware on-device model fine-tuning strategy to adapt to new tasks with an accuracy-latency trade-off, and dynamically updates the model zoo to enhance ELITE. Extensive evaluations on five real-world datasets have been conducted, and the results demonstrate that ELITE consistently outperforms the state-of-art solutions, improving the accuracy by 16.3% on average and reducing the response latency by up to 1.98 times.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- A Two-Stage Data Selection Framework for Data-Efficient Model Training on Edge DevicesChen Gong, Rui Xing, Zhenzhe Zheng, Fan WuKDD 2025 · 被引用 1 次
- Smaller but Better: Plasticity-Preserving Continual Learning for Embedded AIChenxin Mao, Haibo Liu, Zhenzhe Zheng, Fan Wu 等WWW 2026
它引用的顶会 Paper22
- Learning to Prompt for Continual LearningZifeng Wang, Zizhao Zhang, Chen-Yu Lee, Han Zhang 等CVPR 2022 · 被引用 635 次
- Online Class-Incremental Continual Learning with Adversarial Shapley ValueDongsub Shim, Zheda Mai, Jihwan Jeong, Scott Sanner 等AAAI 2021 · 被引用 262 次
- Scalable and Order-robust Continual Learning with Additive Parameter DecompositionJaehong Yoon, Saehoon Kim, Eunho Yang, Sung Ju HwangICLR 2020 · 被引用 206 次
- Online Continual Learning from Imbalanced DataAristotelis Chrysakis, Marie-Francine MoensICML 2020 · 被引用 166 次
- Online Continual Learning through Mutual Information MaximizationYiduo Guo, Bing Liu, Dongyan ZhaoICML 2022 · 被引用 139 次
相关 Paper
- Lethe: Plasticity-aware Active Forgetting for Resource-Efficient On-Device Continual LearningHaibo Liu, Chenxin Mao, Zhenzhe Zheng, Fan Wu 等KDD 2026
- Delta: A Cloud-assisted Data Enrichment Framework for On-Device Continual LearningChen Gong, Zhenzhe Zheng, Fan Wu, Xiaofeng Jia 等MobiCom 2024 · 被引用 6 次
- Latency-Aware Online Continual Learning for Non-Stationary Data StreamsHaibo Liu, Da Huo, Zhenzhe Zheng, Fan WuINFOCOM 2025
- RECL: Responsive Resource-Efficient Continuous Learning for Video AnalyticsMehrdad Khani Shirkoohi, Ganesh Ananthanarayanan, Kevin Hsieh, Junchen Jiang 等NSDI 2023
- DUET: A Tuning-Free Device-Cloud Collaborative Parameters Generation Framework for Efficient Device Model GeneralizationZheqi Lv, Wenqiao Zhang, Shengyu Zhang, Kun Kuang 等WWW 2023 · 被引用 68 次
