Vehicle Counting Network with Attention-based Mask Refinement and Spatial-awareness Block Loss
Ji Zhang, Jian-Jun Qiao, Xiao Wu, Wei Li
Abstract
Vehicle counting aims to calculate the number of vehicles in congested traffic scenes. Although object detection and crowd counting have made tremendous progress with the development of deep learning, vehicle counting remains a challenging task, due to scale variations, viewpoint changes, inconsistent location distributions, diverse visual appearances and severe occlusions. In this paper, a well-designed Vehicle Counting Network (VCNet) is novelly proposed to alleviate the problem of scale variation and inconsistent spatial distribution in congested traffic scenes. Specifically, VCNet is composed of two major components: (i) To capture multi-scale vehicles across different types and camera viewpoints, an effective multi-scale density map estimation structure is designed by building an attention-based mask refinement module. The multi-branch structure with hybrid dilated convolution blocks is proposed to assign receptive fields to generate multi-scale density maps. To efficiently aggregate multi-scale density maps, the attention-based mask refinement is well-designed to highlight the vehicle regions, which enables each branch to suppress the scale interference from other branches. (ii) In order to capture the inconsistent spatial distributions, a spatial-awareness block loss (SBL) based on the region-weighted reward strategy is proposed to calculate the loss of different spatial regions including sparse, congested and occluded regions independently by dividing the density map into different regions. Extensive experiments conducted on three benchmark datasets, TRANCOS, VisDrone2019 Vehicle and CVCSet demonstrate that the proposed VCNet outperforms the state-of-the-art approaches in vehicle counting. Moreover, the proposed idea can be applicable for crowd counting, which produces competitive results on ShanghaiTech crowd counting dataset.
Ask about this paper
Ask your agent about it.
Lune has read the top-tier papers around this one, so every answer names the papers it rests on.
Your agent calls
Lunesearch_papers
Free to start. No credit card required.
Terminal
Install the CLIlune papers get 673087e4-aa79-4237-bcff-d2d6a6735350Related papers
- CODAN: Counting-driven Attention Network for Vehicle Detection in Congested ScenesWei Li, Zhenting Wang, Xiao Wu, Ji Zhang et al.ACM MM 2020 · 19 citations
- Reverse Perspective Network for Perspective-Aware Object CountingYifan Yang, Guorong Li, Zhe Wu, Li Su et al.CVPR 2020
- From Open Set to Closed Set: Counting Objects by Spatial Divide-and-ConquerHaipeng Xiong, Hao Lu, Chengxin Liu, Liang Liu et al.ICCV 2019 · 184 citations
- Attention Scaling for Crowd CountingXiaoheng Jiang, Li Zhang, Mingliang Xu, Tianzhu Zhang et al.CVPR 2020
- Crowd Counting With Deep Structured Scale Integration NetworkLingbo Liu, Zhilin Qiu, Guanbin Li, Shufan Liu et al.ICCV 2019 · 254 citations
