Video Individual Counting for Moving Drones
Yaowu Fan, Jia Wan, Tao Han, Antoni B. Chan, Andy J. Ma
Abstract
Video Individual Counting (VIC) has received increasing attention for its importance in intelligent video surveillance. Existing works are limited in two aspects, i.e., dataset and method. Previous datasets are captured with fixed or rarely moving cameras with relatively sparse individuals, restricting evaluation for a highly varying view and time in crowded scenes. Existing methods rely on localization followed by association or classification, which struggle under dense and dynamic conditions due to inaccurate localization of small targets. To address these issues, we introduce the MovingDroneCrowd Dataset, featuring videos captured by fast-moving drones in crowded scenes under diverse illuminations, shooting heights and angles. We further propose a Shared Density map-guided Network (SDNet) using a Depth-wise Cross-Frame Attention (DCFA) module to directly estimate shared density maps between consecutive frames, from which the inflow and outflow density maps are derived by subtracting the shared density maps from the global density maps. The inflow density maps across frames are summed up to obtain the number of unique pedestrians in a video. Experiments on our datasets and publicly available ones show the superiority of our method over the state of the arts in highly dynamic and complex crowded scenes. Our dataset and codes have been released publicly 1 .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 43c0d9ea-5344-43d1-8d63-682cd561f502Cited by top-tier papers1
Ask how each one uses itBuilds on18
- TrackFormer: Multi-Object Tracking with TransformersTim Meinhardt, Alexander Kirillov, Laura Leal-Taixé, Christoph FeichtenhoferCVPR 2022 · 927 citations
- Rethinking Counting and Localization in Crowds: A Purely Point-Based FrameworkQingyu Song, Changan Wang, Zhengkai Jiang, Yabiao Wang et al.ICCV 2021 · 376 citations
- DanceTrack: Multi-Object Tracking in Uniform Appearance and Diverse MotionPeize Sun, Jinkun Cao, Yi Jiang, Zehuan Yuan et al.CVPR 2022 · 305 citations
- MeMOT: Multi-Object Tracking with MemoryJiarui Cai, Mingze Xu, Wei Li, Yuanjun Xiong et al.CVPR 2022 · 216 citations
- Perspective-Guided Convolution Networks for Crowd CountingZhaoyi Yan, Yuchen Yuan, Wangmeng Zuo, Xiao Tan et al.ICCV 2019 · 209 citations
Related papers
- DR.VIC: Decomposition and Reasoning for Video Individual CountingTao Han, Lei Bai, Junyu Gao, Qi Wang et al.CVPR 2022 · 18 citations
- Weakly Supervised Video Individual CountingXinyan Liu, Guorong Li, Yuankai Qi, Ziheng Yan et al.CVPR 2024
- Flowing Crowd to Count Flows: A Self-Supervised Framework for Video Individual CountingFeng-Kai Huang, Bo-Lun Huang, Li-Wu Tsao, Jhih-Ciang Wu et al.ACM MM 2025 · 1 citation
- Detection, Tracking, and Counting Meets Drones in Crowds: A BenchmarkLongyin Wen, Dawei Du, Pengfei Zhu, Qinghua Hu et al.CVPR 2021
- Efficient Masked AutoEncoder for Video Object Counting and A Large-Scale BenchmarkBing Cao, Quanhao Lu, Jiekang Feng, Qilong Wang et al.ICLR 2025
