Detection in Crowded Scenes: One Proposal, Multiple Predictions
Xuangeng Chu, Anlin Zheng, Xiangyu Zhang, Jian Sun
Abstract
We propose a simple yet effective proposal-based object detector, aiming at detecting highly-overlapped instances in crowded scenes. The key of our approach is to let each proposal predict a set of correlated instances rather than a single one in previous proposalbased frameworks. Equipped with new techniques such as EMD Loss and Set NMS, our detector can effectively handle the difficulty of detecting highly overlapped objects. On a FPN-Res50 baseline, our detector can obtain 4.9% AP gains on challenging CrowdHuman dataset and 1.0% MR -2 improvements on CityPersons dataset, without bells and whistles. Moreover, on less crowed datasets like COCO, our approach can still achieve moderate improvement, suggesting the proposed method is robust to crowdedness. Code and pre-trained models will be released at https://github.com/megvii-model/CrowdDetection.
The work is done when Xuangeng Chu was an intern in MEGVII Technology.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext ecf9ce7b-83c1-49aa-8923-10e3ebba2c2cCited by top-tier papers23
- NDC-Scene: Boost Monocular 3D Semantic Scene Completion in Normalized Device Coordinates SpaceJiawei Yao, Chuming Li, Keqiang Sun, Yingjie Cai et al.ICCV 2023 · 150 citations
- Progressive End-to-End Object Detection in Crowded ScenesAnlin Zheng, Yuang Zhang, Xiangyu Zhang, Xiaojuan Qi et al.CVPR 2022 · 82 citations
- Multi-Instance Pose Networks: Rethinking Top-Down Pose EstimationRawal Khirodkar, Visesh Chari, Amit Agrawal, Ambrish TyagiICCV 2021 · 80 citations
- Amodal Segmentation Based on Visible Region Segmentation and Shape PriorYuting Xiao, Yanyu Xu, Ziming Zhong, Weixin Luo et al.AAAI 2021 · 76 citations
- 4D-Net for Learned Multi-Modal AlignmentA. J. Piergiovanni, Vincent Casser, Michael S. Ryoo, Anelia AngelovaICCV 2021 · 69 citations
Builds on1
Related papers
- NMS by Representative Region: Towards Crowded Pedestrian Detection by Proposal PairingXin Huang, Zheng Ge, Zequn Jie, Osamu YoshieCVPR 2020
- Relational Learning for Joint Head and Human DetectionCheng Chi, Shifeng Zhang, Junliang Xing, Zhen Lei et al.AAAI 2020 · 61 citations
- Body-Face Joint Detection via Embedding and Head HookJunfeng Wan, Jiangfan Deng, Xiaosong Qiu, Feng ZhouICCV 2021 · 17 citations
- Instance Guided Proposal Network for Person SearchWenkai Dong, Zhaoxiang Zhang, Chunfeng Song, Tieniu TanCVPR 2020
- DiffusionDet: Diffusion Model for Object DetectionShoufa Chen, Peize Sun, Yibing Song, Ping LuoICCV 2023 · 715 citations
