KDSelector: A Framework of Knowledge-Enhanced and Data-Efficient Selector Learning for Anomaly Detection Model Selection in Time Series
Zhiyu Liang, Dongrui Cai, Chenyuan Zhang, Zheng Liang, Chen Liang, Shi Qiu, Jin Wang, Hongzhi Wang
Abstract
Model selection has been raised as an essential problem in the area of time series anomaly detection (TSAD), because there is no single best TSAD model for highly heterogeneous time series in real-world applications. However, despite the success of existing model selection solutions, which usually learn (a.k.a. train) a classification model (especially neural network, NN) using historical data as a selector to predict the correct TSAD model for each time series to detect, the existing NN-based selector learning method cannot utilize the auxiliary knowledge in the historical data and requires iterating over all training samples, which limits the model selection ability and training speed of the selector. The latter data efficiency problem can be partially solved by existing data pruning methods designed for general NN training, but with suboptimal speedup or degraded selection ability due to disregarding intrinsic data properties in TSAD model selector training. To address these limitations, we propose KDSelector, to the best of our knowledge, the first framework customized for knowledge-enhanced and data-efficient learning of NN-based TSAD model selectors, of which we design three plug-and-play modules that are agnostic to NN architectures (e.g., ResNet and Transformer) and can be seamlessly integrated into the existing selector learning framework. Specifically, we propose two knowledge enhancement mechanisms to improve the selection ability of the selector with any architecture by integrating the auxiliary knowledge in a unified way. We further design a novel data pruning framework with theoretical guarantees to achieve state-of-the-art training acceleration for the NN-based selector with almost lossless selection ability. Extensive experiments demonstrate the superior performance of our proposals in terms of model selection ability and selector learning efficiency.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Builds on13
- Deep Learning on a Data Diet: Finding Important Examples Early in TrainingMansheej Paul, Surya Ganguli, Gintare Karolina DziugaiteNeurIPS 2021 · 806 citations
- Anomaly Detection in Time Series: A Comprehensive EvaluationSebastian Schmidl, Phillip Wenig, Thorsten PapenbrockVLDB 2022 · 578 citations
- Towards a Rigorous Evaluation of Time-Series Anomaly DetectionSiwon Kim, Kukjin Choi, Hyun-Soo Choi, Byunghan Lee et al.AAAI 2022 · 220 citations
- Volume Under the Surface: A New Accuracy Evaluation Measure for Time-Series Anomaly DetectionJohn Paparrizos, Paul Boniol, Themis Palpanas, Ruey S. Tsay et al.VLDB 2022 · 171 citations
- TSB-UAD: An End-to-End Benchmark Suite for Univariate Time-Series Anomaly DetectionJohn Paparrizos, Yuhao Kang, Paul Boniol, Ruey S. Tsay et al.VLDB 2022 · 138 citations
Related papers
- Choose Wisely: An Extensive Evaluation of Model Selection for Anomaly Detection in Time SeriesEmmanouil Sylligardos, Paul Boniol, John Paparrizos, Panos E. Trahanias et al.VLDB 2023 · 40 citations
- TSB-AutoAD: Towards Automated Solutions for Time-Series Anomaly Detection [E, A & B]Qinghua Liu, Seunghak Lee, John PaparrizosVLDB 2025 · 13 citations
- Evolving Proxy Kills Drift: Data-Efficient Streaming Time Series Anomaly DetectionQing Wei, Hao Miao, Yan Zhao, Kai Zheng et al.WWW 2026 · 1 citation
- KARMAD: KAN-Based Adversarial Robust Model for Anomaly DetectionFangke Chen, Xiaotian Qiu, Yihan Ye, Ruyue Jing et al.ICDE 2025 · 2 citations
- TranAD: Deep Transformer Networks for Anomaly Detection in Multivariate Time Series DataShreshth Tuli, Giuliano Casale, Nicholas R. JenningsVLDB 2022 · 930 citations
