DAVE - A Detect-and-Verify Paradigm for Low-Shot Counting
Jer Pelhan, Alan Lukezic, Vitjan Zavrtanik, Matej Kristan
Abstract
Low-shot counters estimate the number of objects corresponding to a selected category, based on only few or no exemplars annotated in the image. The current state-of-the-art estimates the total counts as the sum over the object location density map, but does not provide individual object locations and sizes, which are crucial for many applications. This is addressed by detection-based counters, which, however fall behind in the total count accuracy. Furthermore, both approaches tend to overestimate the counts in the presence of other object classes due to many false positives. We propose DAVE, a low-shot counter based on a detect-and-verify paradigm, that avoids the aforementioned issues by first generating a high-recall detection set and then verifying the detections to identify and remove the out-liers. This jointly increases the recall and precision, leading to accurate counts. DAVE outperforms the top density-based counters by 20% in the total count MAE, it outper-forms the most recent detection-based counter by 20% in detection quality and sets a new state-of-the-art in zero-shot as well as text-prompt-based counting. The code and models are available on GitHub.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers20
- CountGD: Multi-Modal Open-World CountingNiki Amini-Naieni, Tengda Han, Andrew ZissermanNeurIPS 2024 · 96 citations
- Detect Anything via Next Point PredictionQing Jiang, Junan Huo, Xingyu Chen, Yuda Xiong et al.CVPR 2026 · 79 citations
- A Novel Unified Architecture for Low-Shot Counting by Detection and SegmentationJer Pelhan, Alan Lukezic, Vitjan Zavrtanik, Matej KristanNeurIPS 2024 · 29 citations
- OmniCount: Multi-label Object Counting with Semantic-Geometric PriorsAnindya Mondal, Sauradip Nag, Xiatian Zhu, Anjan DuttaAAAI 2025 · 14 citations
- CountGD++: Generalized Prompting for Open-World CountingNiki Amini-Naieni, Andrew ZissermanCVPR 2026 · 14 citations
Builds on12
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- Distribution Matching for Crowd CountingBoyu Wang, Huidong Liu, Dimitris Samaras, Minh Hoai NguyenNeurIPS 2020 · 443 citations
- Modeling Noisy Annotations for Crowd CountingJia Wan, Antoni B. ChanNeurIPS 2020 · 120 citations
- Rethinking Spatial Invariance of Convolutional Networks for Object CountingZhi-Qi Cheng, Qi Dai, Hong Li, Jingkuan Song et al.CVPR 2022 · 119 citations
- Represent, Compare, and Learn: A Similarity-Aware Framework for Class-Agnostic CountingMin Shi, Hao Lu, Chen Feng, Chengxin Liu et al.CVPR 2022 · 99 citations
Related papers
- Learning To Count EverythingViresh Ranjan, Udbhav Sharma, Thu Nguyen, Minh HoaiCVPR 2021
- Generalized-Scale Object Counting with Gradual Query AggregationJer Pelhan, Alan Lukezic, Matej KristanAAAI 2026
- Point, Segment and Count: A Generalized Framework for Object CountingZhizhong Huang, Mingliang Dai, Yi Zhang, Junping Zhang et al.CVPR 2024
- A Low-Shot Object Counting Network With Iterative Prototype AdaptationNikola Ðukic, Alan Lukezic, Vitjan Zavrtanik, Matej KristanICCV 2023 · 91 citations
- CountSE: Soft Exemplar Open-Set Object CountingShuai Liu, Peng Zhang, Shiwei Zhang, Wei KeICCV 2025 · 1 citation
