Rethinking Spatial Invariance of Convolutional Networks for Object Counting
Zhi-Qi Cheng, Qi Dai, Hong Li, Jingkuan Song, Xiao Wu, Alexander G. Hauptmann
摘要
Previous work generally believes that improving the spatial invariance of convolutional networks is the key to object counting. However, after verifying several mainstream counting networks, we surprisingly found too strict pixel-level spatial invariance would cause overfit noise in the density map generation. In this paper, we try to use locally connected Gaussian kernels to replace the original convolution filter to estimate the spatial position in the density map. The purpose of this is to allow the feature extraction process to potentially stimulate the density map generation process to overcome the annotation noise. Inspired by previous work, we propose a low-rank approximation accompanied with translation invariance to favorably implement the approximation of massive Gaussian convolution. Our work points a new direction for follow-up research, which should investigate how to properly relax the overly strict pixel-level spatial invariance for object counting. We evaluate our methods on 4 mainstream object counting networks (i.e., MCNN, CSRNet, SANet, and ResNet-50). Extensive experiments were conducted on 7 popular benchmarks for 3 applications (i.e., crowd, vehicle, and plant counting). Experimental results show that our methods significantly outperform other state-of-the-art methods and achieve promising learning of the spatial position of objects <sup xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink">1</sup> <sup xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink">1</sup> Code is at https://github.com/zhiqic/Rethinking-Counting.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper16
- Point-Query Quadtree for Crowd Counting, Localization, and MoreChengxin Liu, Hao Lu, Zhiguo Cao, Tongliang LiuICCV 2023 · 被引用 89 次
- STEERER: Resolving Scale Variations for Counting and Localization via Selective Inheritance LearningTao Han, Lei Bai, Lingbo Liu, Wanli OuyangICCV 2023 · 被引用 74 次
- Vision Transformer Off-the-Shelf: A Surprising Baseline for Few-Shot Class-Agnostic CountingZhicheng Wang, Liwen Xiao, Zhiguo Cao, Hao LuAAAI 2024 · 被引用 35 次
- PoSynDA: Multi-Hypothesis Pose Synthesis Domain Adaptation for Robust 3D Human Pose EstimationHanbing Liu, Jun-Yan He, Zhi-Qi Cheng, Wangmeng Xiang 等ACM MM 2023 · 被引用 30 次
- Counting Crowds in Bad WeatherZhi-Kai Huang, Wei-Ting Chen, Yuan-Chun Chiang, Sy-Yen Kuo 等ICCV 2023 · 被引用 23 次
它引用的顶会 Paper26
- Bayesian Loss for Crowd Count Estimation With Point SupervisionZhiheng Ma, Xing Wei, Xiaopeng Hong, Yihong GongICCV 2019 · 被引用 612 次
- Distribution Matching for Crowd CountingBoyu Wang, Huidong Liu, Dimitris Samaras, Minh Hoai NguyenNeurIPS 2020 · 被引用 443 次
- Rethinking Counting and Localization in Crowds: A Purely Point-Based FrameworkQingyu Song, Changan Wang, Zhengkai Jiang, Yabiao Wang 等ICCV 2021 · 被引用 376 次
- To Choose or to Fuse? Scale Selection for Crowd CountingQingyu Song, Changan Wang, Yabiao Wang, Ying Tai 等AAAI 2021 · 被引用 195 次
- From Open Set to Closed Set: Counting Objects by Spatial Divide-and-ConquerHaipeng Xiong, Hao Lu, Chengxin Liu, Liang Liu 等ICCV 2019 · 被引用 184 次
相关 Paper
- Modeling Noisy Annotations for Crowd CountingJia Wan, Antoni B. ChanNeurIPS 2020 · 被引用 120 次
- Revisiting Spatial Invariance with Low-Rank Local ConnectivityGamaleldin F. Elsayed, Prajit Ramachandran, Jonathon Shlens, Simon KornblithICML 2020 · 被引用 51 次
- Vehicle Counting Network with Attention-based Mask Refinement and Spatial-awareness Block LossJi Zhang, Jian-Jun Qiao, Xiao Wu, Wei LiACM MM 2021 · 被引用 3 次
- Adaptive Density Map Generation for Crowd CountingJia Wan, Antoni B. ChanICCV 2019 · 被引用 171 次
- Learning Spatial Awareness to Improve Crowd CountingZhi-Qi Cheng, Jun-Xiu Li, Qi Dai, Xiao Wu 等ICCV 2019 · 被引用 139 次
