Cross-View Cross-Scene Multi-View Crowd Counting
Qi Zhang, Wei Lin, Antoni B. Chan
Abstract
Multi-view crowd counting has been previously proposed to utilize multi-cameras to extend the field-of-view of a single camera, capturing more people in the scene, and improve counting performance for occluded people or those in low resolution. However, the current multi-view paradigm trains and tests on the same single scene and camera-views, which limits its practical application. In this paper, we propose a cross-view cross-scene (CVCS) multi-view crowd counting paradigm, where the training and testing occur on different scenes with arbitrary camera layouts. To dynamically handle the challenge of optimal view fusion under scene and camera layout change and non-correspondence noise due to camera calibration errors or erroneous features, we propose a CVCS model that attentively selects and fuses multiple views together using camera layout geometry, and a noise view regularization method to train the model to handle non-correspondence errors. We also generate a large synthetic multi-camera crowd counting dataset with a large number of scenes and camera views to capture many possible variations, which avoids the difficulty of collecting and annotating such a large real dataset. We then test our trained CVCS model on real multi-view counting datasets, by using unsupervised domain transfer. The proposed CVCS model trained on synthetic data outperforms the same model trained only on real data, and achieves promising performance compared to fully supervised methods that train and test on the same single scene.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 43a95d50-942f-4419-b7fd-d6ef9e869d3dCited by top-tier papers12
- Rethinking Spatial Invariance of Convolutional Networks for Object CountingZhi-Qi Cheng, Qi Dai, Hong Li, Jingkuan Song et al.CVPR 2022 · 119 citations
- Stacked Homography Transformations for Multi-View Pedestrian DetectionLiangchen Song, Jialian Wu, Ming Yang, Qian Zhang et al.ICCV 2021 · 66 citations
- 3D Crowd Counting via Multi-View Fusion with 3D Gaussian KernelsQi Zhang, Antoni B. ChanAAAI 2020 · 41 citations
- DR.VIC: Decomposition and Reasoning for Video Individual CountingTao Han, Lei Bai, Junyu Gao, Qi Wang et al.CVPR 2022 · 18 citations
- A Fixed-Point Approach to Unified Prompt-Based CountingWei Lin, Antoni B. ChanAAAI 2024 · 11 citations
Builds on9
- Habitat: A Platform for Embodied AI ResearchManolis Savva, Jitendra Malik, Devi Parikh, Dhruv Batra et al.ICCV 2019 · 1,863 citations
- Bayesian Loss for Crowd Count Estimation With Point SupervisionZhiheng Ma, Xing Wei, Xiaopeng Hong, Yihong GongICCV 2019 · 612 citations
- Learnable Triangulation of Human PoseKarim Iskakov, Egor Burkov, Victor S. Lempitsky, Yury MalkovICCV 2019 · 419 citations
- Adaptive Density Map Generation for Crowd CountingJia Wan, Antoni B. ChanICCV 2019 · 171 citations
- Pushing the Frontiers of Unconstrained Crowd Counting: New Dataset and Benchmark MethodVishwanath Sindagi, Rajeev Yasarla, Vishal M. PatelICCV 2019 · 101 citations
Related papers
- Leveraging Self-Supervision for Cross-Domain Crowd CountingWeizhe Liu, Nikita Durasov, Pascal FuaCVPR 2022 · 43 citations
- Multi-View People Detection in Large Scenes via Supervised View-Wise Contribution WeightingQi Zhang, Yunfei Gong, Daijie Chen, Antoni B. Chan et al.AAAI 2024 · 7 citations
- Multi-view Crowd Tracking Transformer with View-Ground Interactions Under Large Real-World ScenesQi Zhang, Jixuan Chen, Zhang Kaiyi, Xinquan Yu et al.CVPR 2026
- Fine-Grained Fragment Diffusion for Cross Domain Crowd CountingHuilin Zhu, Jingling Yuan, Zhengwei Yang, Xian Zhong et al.ACM MM 2022 · 28 citations
- Dynamic Momentum Adaptation for Zero-Shot Cross-Domain Crowd CountingQiangqiang Wu, Jia Wan, Antoni B. ChanACM MM 2021 · 34 citations
