Vehicle Counting Network with Attention-based Mask Refinement and Spatial-awareness Block Loss
Ji Zhang, Jian-Jun Qiao, Xiao Wu, Wei Li
摘要
Vehicle counting aims to calculate the number of vehicles in congested traffic scenes. Although object detection and crowd counting have made tremendous progress with the development of deep learning, vehicle counting remains a challenging task, due to scale variations, viewpoint changes, inconsistent location distributions, diverse visual appearances and severe occlusions. In this paper, a well-designed Vehicle Counting Network (VCNet) is novelly proposed to alleviate the problem of scale variation and inconsistent spatial distribution in congested traffic scenes. Specifically, VCNet is composed of two major components: (i) To capture multi-scale vehicles across different types and camera viewpoints, an effective multi-scale density map estimation structure is designed by building an attention-based mask refinement module. The multi-branch structure with hybrid dilated convolution blocks is proposed to assign receptive fields to generate multi-scale density maps. To efficiently aggregate multi-scale density maps, the attention-based mask refinement is well-designed to highlight the vehicle regions, which enables each branch to suppress the scale interference from other branches. (ii) In order to capture the inconsistent spatial distributions, a spatial-awareness block loss (SBL) based on the region-weighted reward strategy is proposed to calculate the loss of different spatial regions including sparse, congested and occluded regions independently by dividing the density map into different regions. Extensive experiments conducted on three benchmark datasets, TRANCOS, VisDrone2019 Vehicle and CVCSet demonstrate that the proposed VCNet outperforms the state-of-the-art approaches in vehicle counting. Moreover, the proposed idea can be applicable for crowd counting, which produces competitive results on ShanghaiTech crowd counting dataset.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
相关 Paper
- CODAN: Counting-driven Attention Network for Vehicle Detection in Congested ScenesWei Li, Zhenting Wang, Xiao Wu, Ji Zhang 等ACM MM 2020 · 被引用 19 次
- Reverse Perspective Network for Perspective-Aware Object CountingYifan Yang, Guorong Li, Zhe Wu, Li Su 等CVPR 2020
- From Open Set to Closed Set: Counting Objects by Spatial Divide-and-ConquerHaipeng Xiong, Hao Lu, Chengxin Liu, Liang Liu 等ICCV 2019 · 被引用 184 次
- Attention Scaling for Crowd CountingXiaoheng Jiang, Li Zhang, Mingliang Xu, Tianzhu Zhang 等CVPR 2020
- Crowd Counting With Deep Structured Scale Integration NetworkLingbo Liu, Zhilin Qiu, Guanbin Li, Shufan Liu 等ICCV 2019 · 被引用 254 次
