Volume Under the Surface: A New Accuracy Evaluation Measure for Time-Series Anomaly Detection
John Paparrizos, Paul Boniol, Themis Palpanas, Ruey S. Tsay, Aaron J. Elmore, Michael J. Franklin
摘要
Anomaly detection (AD) is a fundamental task for time-series analytics with important implications for the downstream performance of many applications. In contrast to other domains where AD mainly focuses on point-based anomalies (i.e., outliers in standalone observations), AD for time series is also concerned with range-based anomalies (i.e., outliers spanning multiple observations). Nevertheless, it is common to use traditional point-based information retrieval measures, such as Precision, Recall, and F-score, to assess the quality of methods by thresholding the anomaly score to mark each point as an anomaly or not. However, mapping discrete labels into continuous data introduces unavoidable shortcomings, complicating the evaluation of range-based anomalies. Notably, the choice of evaluation measure may significantly bias the experimental outcome. Despite over six decades of attention, there has never been a large-scale systematic quantitative and qualitative analysis of time-series AD evaluation measures. This paper extensively evaluates quality measures for time-series AD to assess their robustness under noise, misalignments, and different anomaly cardinality ratios. Our results indicate that measures producing quality values independently of a threshold (i.e., AUC-ROC and AUC-PR) are more suitable for time-series AD. Motivated by this observation, we first extend the AUC-based measures to account for range-based anomalies. Then, we introduce a new family of parameter-free and threshold-independent measures, VUS (Volume Under the Surface), to evaluate methods while varying parameters. Our findings demonstrate that our four measures are significantly more robust in assessing the quality of time-series AD methods.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper59
- MOMENT: A Family of Open Time-series Foundation ModelsMononito Goswami, Konrad Szafer, Arjun Choudhry, Yifu Cai 等ICML 2024 · 被引用 442 次
- DCdetector: Dual Attention Contrastive Representation Learning for Time Series Anomaly DetectionYiyuan Yang, Chaoli Zhang, Tian Zhou, Qingsong Wen 等KDD 2023 · 被引用 244 次
- TSB-UAD: An End-to-End Benchmark Suite for Univariate Time-Series Anomaly DetectionJohn Paparrizos, Yuhao Kang, Paul Boniol, Ruey S. Tsay 等VLDB 2022 · 被引用 138 次
- ImDiffusion: Imputed Diffusion Models for Multivariate Time Series Anomaly DetectionYuhang Chen, Chaoyun Zhang, Minghua Ma, Yudong Liu 等VLDB 2024 · 被引用 122 次
- Elpis: Graph-Based Similarity Search for Scalable Data ScienceIlias Azizi, Karima Echihabi, Themis PalpanasVLDB 2023 · 被引用 67 次
它引用的顶会 Paper8
- TSB-UAD: An End-to-End Benchmark Suite for Univariate Time-Series Anomaly DetectionJohn Paparrizos, Yuhao Kang, Paul Boniol, Ruey S. Tsay 等VLDB 2022 · 被引用 138 次
- SAND: Streaming Subsequence Anomaly DetectionPaul Boniol, John Paparrizos, Themis Palpanas, Michael J. FranklinVLDB 2021 · 被引用 128 次
- Decomposed Bounded Floats for Fast Compression and QueriesChunwei Liu, Hao Jiang, John Paparrizos, Aaron J. ElmoreVLDB 2021 · 被引用 65 次
- Debunking Four Long-Standing Misconceptions of Time-Series Distance MeasuresJohn Paparrizos, Chunwei Liu, Aaron J. Elmore, Michael J. FranklinSIGMOD 2020 · 被引用 56 次
- Good to the Last Bit: Data-Driven Encoding with CodecDBHao Jiang, Chunwei Liu, John Paparrizos, Andrew A. Chien 等SIGMOD 2021 · 被引用 45 次
相关 Paper
- Local Evaluation of Time Series Anomaly Detection AlgorithmsAlexis Huet, José Manuel Navarro, Dario RossiKDD 2022 · 被引用 73 次
- PATE: Proximity-Aware Time Series Anomaly EvaluationRamin Ghorbani, Marcel J. T. Reinders, David M. J. TaxKDD 2024 · 被引用 11 次
- TAB: Unified Benchmarking of Time Series Anomaly Detection MethodsXiangfei Qiu, Zhe Li, Wanghui Qiu, Shiyan Hu 等VLDB 2025 · 被引用 57 次
- Unsupervised Model Selection for Time Series Anomaly DetectionMononito Goswami, Cristian I. Challu, Laurent Callot, Lenon Minorics 等ICLR 2023 · 被引用 8 次
- Toward Interpretable Evaluation Measures for Time Series SegmentationFélix Chavelli, Paul Boniol, Michaël ThomazoNeurIPS 2025 · 被引用 1 次
