QCore: Data-Efficient, On-Device Continual Calibration for Quantized Models
David Campos, Bin Yang, Tung Kieu, Miao Zhang, Chenjuan Guo, Christian S. Jensen
摘要
We are witnessing an increasing availability of streaming data that may contain valuable information on the underlying processes. It is thus attractive to be able to deploy machine learning models, e.g., for classification, on edge devices near sensors such that decisions can be made instantaneously, rather than first having to transmit incoming data to servers. To enable deployment on edge devices with limited storage and computational capabilities, the full-precision parameters in standard models can be quantized to use fewer bits. The resulting quantized models are then calibrated using back-propagation with the full training data to ensure accuracy. This one-time calibration works for deployments in static environments. However, model deployment in dynamic edge environments call for continual calibration to adaptively adjust quantized models to fit new incoming data, which may have different distributions with the original training data. The first difficulty in enabling continual calibration on the edge is that the full training data may be too large and thus cannot be assumed to be always available on edge devices. The second difficulty is that the use of back-propagation on the edge for repeated calibration is too expensive. We propose QCore to enable continual calibration on the edge. First, it compresses the full training data into a small subset to enable effective calibration of quantized models with different bit-widths. We also propose means of updating the subset when new streaming data arrives to reflect changes in the environment, while not forgetting earlier training data. Second, we propose a small bit-flipping network that works with the subset to update quantized model parameters, thus enabling efficient continual calibration without back-propagation. An experimental study, conducted with real-world data in a continual learning setting, offers insight into the properties of QCore and shows that it is capable of outperforming strong baseline methods.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper7
- Fully Automated Correlated Time Series Forecasting in MinutesXinle Wu, Xingjian Wu, Dalin Zhang, Miao Zhang 等VLDB 2025 · 被引用 17 次
- TEAM: Topological Evolution-aware Framework for Traffic ForecastingDuc Kieu, Tung Kieu, Peng Han, Bin Yang 等VLDB 2025 · 被引用 13 次
- Towards Lightweight Time Series Forecasting: A Patch-Wise Transformer with Weak Data EnrichingMeng Wang, Jintao Yang, Bin Yang, Hui Li 等ICDE 2025 · 被引用 10 次
- Noise Matters: Cross Contrastive Learning for Flink Anomaly DetectionZhihao Zhuang, Yingying Zhang, Kai Zhao, Chenjuan Guo 等VLDB 2025 · 被引用 3 次
- A Memory Guided Transformer for Time Series ForecastingYunyao Cheng, Chenjuan Guo, Bin Yang, Haomin Yu 等VLDB 2025 · 被引用 1 次
它引用的顶会 Paper34
- Dark Experience for General Continual Learning: a Strong, Simple BaselinePietro Buzzega, Matteo Boschini, Angelo Porrello, Davide Abati 等NeurIPS 2020 · 被引用 1,494 次
- MCUNet: Tiny Deep Learning on IoT DevicesJi Lin, Wei-Ming Chen, Yujun Lin, John Cohn 等NeurIPS 2020 · 被引用 827 次
- Up or Down? Adaptive Rounding for Post-Training QuantizationMarkus Nagel, Rana Ali Amjad, Mart van Baalen, Christos Louizos 等ICML 2020 · 被引用 816 次
- Coresets for Data-efficient Training of Machine Learning ModelsBaharan Mirzasoleiman, Jeff A. Bilmes, Jure LeskovecICML 2020 · 被引用 494 次
- MiniRocket: A Very Fast (Almost) Deterministic Transform for Time Series ClassificationAngus Dempster, Daniel F. Schmidt, Geoffrey I. WebbKDD 2021 · 被引用 395 次
相关 Paper
- No Retraining at Edge: Efficient Resource-Aware Mixed-Precision Quantization via Federated Supernet LearningLianbo Ma, Yonghui Su, Nan Li, Xingwei WangICML 2026
- Resource-Constrained Federated Continual Learning: What Does Matter?Yichen Li, Yuying Wang, Jiahua Dong, Haozhao Wang 等NeurIPS 2025 · 被引用 7 次
- LEAF: An Adaptation Framework against Noisy Data on Edge through Ultra Low-Cost TrainingZihan Xia, Jinwook Kim, Mingu KangDAC 2024
- Octo: INT8 Training with Loss-aware Compensation and Backward Quantization for Tiny On-device LearningQihua Zhou, Song Guo, Zhihao Qu, Jingcai Guo 等USENIX ATC 2021 · 被引用 55 次
- Robust Machine Unlearning for Quantized Neural Networks via Adaptive Gradient Reweighting with Similar LabelsYujia Tong, Yuze Wang, Jingling Yuan, Chuang HuICCV 2025 · 被引用 3 次
