A Delay Metric for Video Object Detection: What Average Precision Fails to Tell
Huizi Mao, Xiaodong Yang, Bill Dally
摘要
Average precision (AP) is a widely used metric to evaluate detection accuracy of image and video object detectors. In this paper, we analyze object detection from videos and point out that AP alone is not sufficient to capture the temporal nature of video object detection. To tackle this problem, we propose a comprehensive metric, average delay (AD), to measure and compare detection delay. To facilitate delay evaluation, we carefully select a subset of ImageNet VID, which we name as ImageNet VIDT with an emphasis on complex trajectories. By extensively evaluating a wide range of detectors on VIDT, we show that most methods drastically increase the detection delay but still preserve AP well. In other words, AP is not sensitive enough to reflect the temporal characteristics of a video object detector. Our results suggest that video object detection methods should be additionally evaluated with a delay metric, particularly for latency-critical applications such as autonomous vehicle perception.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- VmAP: A Fair Metric for Video Object DetectionAnupam Sobti, Vaibhav Mavi, M. Balakrishnan, Chetan AroraACM MM 2021 · 被引用 5 次
- Enabling SLO-Aware 5G Multi-Access Edge Computing with SMECXiao Zhang, Daehyeok KimNSDI 2026 · 被引用 4 次
- Perception Characteristics Distance: Measuring Stability and Robustness of Perception System in Dynamic Conditions under a Certain Decision RuleBoyu Jiang, Liang Shi, Zhengzhi Lin, Lanxin Xiang 等CVPR 2026 · 被引用 1 次
- Learning to Evaluate Perception Models Using Planner-Centric MetricsJonah Philion, Amlan Kar, Sanja FidlerCVPR 2020
它引用的顶会 Paper1
相关 Paper
- Temporal Context Enhanced Feature Aggregation for Video Object DetectionFei He, Naiyu Gao, Qiaozhe Li, Senyao Du 等AAAI 2020 · 被引用 40 次
- Leveraging Long-Range Temporal Relationships Between Proposals for Video Object DetectionMykhailo Shvets, Wei Liu, Alexander C. BergICCV 2019 · 被引用 91 次
- Feature Aggregated Queries for Transformer-Based Video Object DetectorsYiming CuiCVPR 2023
- Objects do not disappear: Video object detection by single-frame object location anticipationXin Liu, Fatemeh Karimi Nejadasl, Jan C. van Gemert, Olaf Booij 等ICCV 2023 · 被引用 10 次
- QueryProp: Object Query Propagation for High-Performance Video Object DetectionFei He, Naiyu Gao, Jian Jia, Xin Zhao 等AAAI 2022 · 被引用 35 次
