DR.VIC: Decomposition and Reasoning for Video Individual Counting
Tao Han, Lei Bai, Junyu Gao, Qi Wang, Wanli Ouyang
Abstract
Pedestrian counting is a fundamental tool for under-standing pedestrian patterns and crowd flow analysis. Existing works (e.g., image-level pedestrian counting, cross-line crowd counting et al.) either only focus on the image-level counting or are constrained to the manual annotation of lines. In this work, we propose to conduct the pedes-trian counting from a new perspective - Video Individual Counting (VIC), which counts the total number of individual pedestrians in the given video (a person is only counted once). Instead of relying on the Multiple Object Tracking (MOT) techniques, we propose to solve the problem by decomposing all pedestrians into the initial pedestrians who existed in the first frame and the new pedestrians with separate identities in each following frame. Then, an end-to-end Decomposition and Reasoning Network (DRNet) is designed to predict the initial pedestrian count with the density estimation method and reason the new pedestrian's count of each frame with the differentiable optimal transport. Extensive experiments are conducted on two datasets with congested pedestrians and diverse scenes, demonstrating the effectiveness of our method over baselines with great superiority in counting the individual pedestrians. Code: https://github.com/taohan10200/DRNet.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 0608b066-4110-44cd-aef7-07d94550810fCited by top-tier papers11
- CLIP-Count: Towards Text-Guided Zero-Shot Object CountingRuixiang Jiang, Lingbo Liu, Changwen ChenACM MM 2023 · 78 citations
- STEERER: Resolving Scale Variations for Counting and Localization via Selective Inheritance LearningTao Han, Lei Bai, Lingbo Liu, Wanli OuyangICCV 2023 · 74 citations
- Depth-Aware Concealed Crop Detection in Dense Agricultural ScenesLiqiong Wang, Jinyu Yang, Yanfu Zhang, Fangyi Wang et al.CVPR 2024 · 66 citations
- DAOT: Domain-Agnostically Aligned Optimal Transport for Domain-Adaptive Crowd CountingHuilin Zhu, Jingling Yuan, Xian Zhong, Zhengwei Yang et al.ACM MM 2023 · 27 citations
- KITS: Inductive Spatio-Temporal Kriging with Increment Training StrategyQianxiong Xu, Cheng Long, Ziyue Li, Sijie Ruan et al.AAAI 2025 · 19 citations
Builds on11
- Bayesian Loss for Crowd Count Estimation With Point SupervisionZhiheng Ma, Xing Wei, Xiaopeng Hong, Yihong GongICCV 2019 · 612 citations
- Distribution Matching for Crowd CountingBoyu Wang, Huidong Liu, Dimitris Samaras, Minh Hoai NguyenNeurIPS 2020 · 443 citations
- Rethinking Counting and Localization in Crowds: A Purely Point-Based FrameworkQingyu Song, Changan Wang, Zhengkai Jiang, Yabiao Wang et al.ICCV 2021 · 376 citations
- Crowd Counting With Deep Structured Scale Integration NetworkLingbo Liu, Zhilin Qiu, Guanbin Li, Shufan Liu et al.ICCV 2019 · 254 citations
- Localization in the Crowd with Topological ConstraintsShahira Abousamra, Minh Hoai, Dimitris Samaras, Chao ChenAAAI 2021 · 160 citations
Related papers
- Video Individual Counting for Moving DronesYaowu Fan, Jia Wan, Tao Han, Antoni B. Chan et al.ICCV 2025 · 1 citation
- Prototype-Guided Dual-Transformer Reasoning for Video Individual CountingRui Li, Yishu Liu, Huafeng Li, Jinxing Li et al.ACM MM 2024 · 2 citations
- Weakly Supervised Video Individual CountingXinyan Liu, Guorong Li, Yuankai Qi, Ziheng Yan et al.CVPR 2024
- From Open Set to Closed Set: Counting Objects by Spatial Divide-and-ConquerHaipeng Xiong, Hao Lu, Chengxin Liu, Liang Liu et al.ICCV 2019 · 184 citations
- Flowing Crowd to Count Flows: A Self-Supervised Framework for Video Individual CountingFeng-Kai Huang, Bo-Lun Huang, Li-Wu Tsao, Jhih-Ciang Wu et al.ACM MM 2025 · 1 citation
