QCore: Data-Efficient, On-Device Continual Calibration for Quantized Models
David Campos, Bin Yang, Tung Kieu, Miao Zhang, Chenjuan Guo, Christian S. Jensen
Abstract
We are witnessing an increasing availability of streaming data that may contain valuable information on the underlying processes. It is thus attractive to be able to deploy machine learning models, e.g., for classification, on edge devices near sensors such that decisions can be made instantaneously, rather than first having to transmit incoming data to servers. To enable deployment on edge devices with limited storage and computational capabilities, the full-precision parameters in standard models can be quantized to use fewer bits. The resulting quantized models are then calibrated using back-propagation with the full training data to ensure accuracy. This one-time calibration works for deployments in static environments. However, model deployment in dynamic edge environments call for continual calibration to adaptively adjust quantized models to fit new incoming data, which may have different distributions with the original training data. The first difficulty in enabling continual calibration on the edge is that the full training data may be too large and thus cannot be assumed to be always available on edge devices. The second difficulty is that the use of back-propagation on the edge for repeated calibration is too expensive. We propose QCore to enable continual calibration on the edge. First, it compresses the full training data into a small subset to enable effective calibration of quantized models with different bit-widths. We also propose means of updating the subset when new streaming data arrives to reflect changes in the environment, while not forgetting earlier training data. Second, we propose a small bit-flipping network that works with the subset to update quantized model parameters, thus enabling efficient continual calibration without back-propagation. An experimental study, conducted with real-world data in a continual learning setting, offers insight into the properties of QCore and shows that it is capable of outperforming strong baseline methods.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 45664c54-7c4f-445a-9639-b40fa43bf86bCited by top-tier papers7
- Fully Automated Correlated Time Series Forecasting in MinutesXinle Wu, Xingjian Wu, Dalin Zhang, Miao Zhang et al.VLDB 2025 · 17 citations
- TEAM: Topological Evolution-aware Framework for Traffic ForecastingDuc Kieu, Tung Kieu, Peng Han, Bin Yang et al.VLDB 2025 · 13 citations
- Towards Lightweight Time Series Forecasting: A Patch-Wise Transformer with Weak Data EnrichingMeng Wang, Jintao Yang, Bin Yang, Hui Li et al.ICDE 2025 · 10 citations
- Noise Matters: Cross Contrastive Learning for Flink Anomaly DetectionZhihao Zhuang, Yingying Zhang, Kai Zhao, Chenjuan Guo et al.VLDB 2025 · 3 citations
- A Memory Guided Transformer for Time Series ForecastingYunyao Cheng, Chenjuan Guo, Bin Yang, Haomin Yu et al.VLDB 2025 · 1 citation
Builds on34
- Dark Experience for General Continual Learning: a Strong, Simple BaselinePietro Buzzega, Matteo Boschini, Angelo Porrello, Davide Abati et al.NeurIPS 2020 · 1,494 citations
- MCUNet: Tiny Deep Learning on IoT DevicesJi Lin, Wei-Ming Chen, Yujun Lin, John Cohn et al.NeurIPS 2020 · 827 citations
- Up or Down? Adaptive Rounding for Post-Training QuantizationMarkus Nagel, Rana Ali Amjad, Mart van Baalen, Christos Louizos et al.ICML 2020 · 816 citations
- Coresets for Data-efficient Training of Machine Learning ModelsBaharan Mirzasoleiman, Jeff A. Bilmes, Jure LeskovecICML 2020 · 494 citations
- MiniRocket: A Very Fast (Almost) Deterministic Transform for Time Series ClassificationAngus Dempster, Daniel F. Schmidt, Geoffrey I. WebbKDD 2021 · 395 citations
Related papers
- No Retraining at Edge: Efficient Resource-Aware Mixed-Precision Quantization via Federated Supernet LearningLianbo Ma, Yonghui Su, Nan Li, Xingwei WangICML 2026
- Resource-Constrained Federated Continual Learning: What Does Matter?Yichen Li, Yuying Wang, Jiahua Dong, Haozhao Wang et al.NeurIPS 2025 · 7 citations
- LEAF: An Adaptation Framework against Noisy Data on Edge through Ultra Low-Cost TrainingZihan Xia, Jinwook Kim, Mingu KangDAC 2024
- Octo: INT8 Training with Loss-aware Compensation and Backward Quantization for Tiny On-device LearningQihua Zhou, Song Guo, Zhihao Qu, Jingcai Guo et al.USENIX ATC 2021 · 55 citations
- Robust Machine Unlearning for Quantized Neural Networks via Adaptive Gradient Reweighting with Similar LabelsYujia Tong, Yuze Wang, Jingling Yuan, Chuang HuICCV 2025 · 3 citations
