DR.VIC: Decomposition and Reasoning for Video Individual Counting
Tao Han, Lei Bai, Junyu Gao, Qi Wang, Wanli Ouyang
摘要
Pedestrian counting is a fundamental tool for under-standing pedestrian patterns and crowd flow analysis. Existing works (e.g., image-level pedestrian counting, cross-line crowd counting et al.) either only focus on the image-level counting or are constrained to the manual annotation of lines. In this work, we propose to conduct the pedes-trian counting from a new perspective - Video Individual Counting (VIC), which counts the total number of individual pedestrians in the given video (a person is only counted once). Instead of relying on the Multiple Object Tracking (MOT) techniques, we propose to solve the problem by decomposing all pedestrians into the initial pedestrians who existed in the first frame and the new pedestrians with separate identities in each following frame. Then, an end-to-end Decomposition and Reasoning Network (DRNet) is designed to predict the initial pedestrian count with the density estimation method and reason the new pedestrian's count of each frame with the differentiable optimal transport. Extensive experiments are conducted on two datasets with congested pedestrians and diverse scenes, demonstrating the effectiveness of our method over baselines with great superiority in counting the individual pedestrians. Code: https://github.com/taohan10200/DRNet.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper11
- CLIP-Count: Towards Text-Guided Zero-Shot Object CountingRuixiang Jiang, Lingbo Liu, Changwen ChenACM MM 2023 · 被引用 78 次
- STEERER: Resolving Scale Variations for Counting and Localization via Selective Inheritance LearningTao Han, Lei Bai, Lingbo Liu, Wanli OuyangICCV 2023 · 被引用 74 次
- Depth-Aware Concealed Crop Detection in Dense Agricultural ScenesLiqiong Wang, Jinyu Yang, Yanfu Zhang, Fangyi Wang 等CVPR 2024 · 被引用 66 次
- DAOT: Domain-Agnostically Aligned Optimal Transport for Domain-Adaptive Crowd CountingHuilin Zhu, Jingling Yuan, Xian Zhong, Zhengwei Yang 等ACM MM 2023 · 被引用 27 次
- KITS: Inductive Spatio-Temporal Kriging with Increment Training StrategyQianxiong Xu, Cheng Long, Ziyue Li, Sijie Ruan 等AAAI 2025 · 被引用 19 次
它引用的顶会 Paper11
- Bayesian Loss for Crowd Count Estimation With Point SupervisionZhiheng Ma, Xing Wei, Xiaopeng Hong, Yihong GongICCV 2019 · 被引用 612 次
- Distribution Matching for Crowd CountingBoyu Wang, Huidong Liu, Dimitris Samaras, Minh Hoai NguyenNeurIPS 2020 · 被引用 443 次
- Rethinking Counting and Localization in Crowds: A Purely Point-Based FrameworkQingyu Song, Changan Wang, Zhengkai Jiang, Yabiao Wang 等ICCV 2021 · 被引用 376 次
- Crowd Counting With Deep Structured Scale Integration NetworkLingbo Liu, Zhilin Qiu, Guanbin Li, Shufan Liu 等ICCV 2019 · 被引用 254 次
- Localization in the Crowd with Topological ConstraintsShahira Abousamra, Minh Hoai, Dimitris Samaras, Chao ChenAAAI 2021 · 被引用 160 次
相关 Paper
- Video Individual Counting for Moving DronesYaowu Fan, Jia Wan, Tao Han, Antoni B. Chan 等ICCV 2025 · 被引用 1 次
- Prototype-Guided Dual-Transformer Reasoning for Video Individual CountingRui Li, Yishu Liu, Huafeng Li, Jinxing Li 等ACM MM 2024 · 被引用 2 次
- Weakly Supervised Video Individual CountingXinyan Liu, Guorong Li, Yuankai Qi, Ziheng Yan 等CVPR 2024
- From Open Set to Closed Set: Counting Objects by Spatial Divide-and-ConquerHaipeng Xiong, Hao Lu, Chengxin Liu, Liang Liu 等ICCV 2019 · 被引用 184 次
- Flowing Crowd to Count Flows: A Self-Supervised Framework for Video Individual CountingFeng-Kai Huang, Bo-Lun Huang, Li-Wu Tsao, Jhih-Ciang Wu 等ACM MM 2025 · 被引用 1 次
