Progressive End-to-End Object Detection in Crowded Scenes
Anlin Zheng, Yuang Zhang, Xiangyu Zhang, Xiaojuan Qi, Jian Sun
摘要
In this paper, we propose a new query-based detection framework for crowd detection. Previous query-based detectors suffer from two drawbacks: first, multiple predictions will be inferred for a single object, typically in crowded scenes; second, the performance saturates as the depth of the decoding stage increases. Benefiting from the nature of the one-to-one label assignment rule, we propose a progressive predicting method to address the above issues. Specifically, we first select accepted queries prone to generate true positive predictions, then refine the rest noisy queries according to the previously accepted predictions. Experiments show that our method can significantly boost the performance of query-based detectors in crowded scenes. Equipped with our approach, Sparse RCNN achieves 92.0% AP, 41.4% MR <sup xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink">−2</sup> and 83.2% JI on the challenging CrowdHuman [35] dataset, outperforming the box-based method MIP [8] that specifies in handling crowded scenarios. Moreover, the proposed method, robust to crowdedness, can still obtain consistent improvements on moderately and slightly crowded datasets like CityPersons [47] and COCO [26]. Code will be made publicly available at https://github.com/megvii-model/Iter-E2EDET.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper11
- Reconstructing Groups of People with Hypergraph Relational ReasoningBuzhen Huang, Jingyi Ju, Zhihao Li, Yangang WangICCV 2023 · 被引用 21 次
- Improving Crowded Object Detection via Copy-PasteJiangfan Deng, Dewen Fan, Xiaosong Qiu, Feng ZhouAAAI 2023 · 被引用 16 次
- RecursiveDet: End-to-End Region-based Recursive Object DetectionJing Zhao, Li Sun, Qingli LiICCV 2023 · 被引用 4 次
- PBADet: A One-Stage Anchor-Free Approach for Part-Body AssociationZhongpai Gao, Huayi Zhou, Abhishek Sharma, Meng Zheng 等ICLR 2024 · 被引用 2 次
- Learning from Synchronization: Self-Supervised Uncalibrated Multi-View Person Association in Challenging ScenesKeqi Chen, Vinkle Srivastav, Didier Mutter, Nicolas PadoyCVPR 2025
它引用的顶会 Paper12
- Deformable DETR: Deformable Transformers for End-to-End Object DetectionXizhou Zhu, Weijie Su, Lewei Lu, Bin Li 等ICLR 2021 · 被引用 7,353 次
- FCOS: Fully Convolutional One-Stage Object DetectionZhi Tian, Chunhua Shen, Hao Chen, Tong HeICCV 2019 · 被引用 6,042 次
- Fast Convergence of DETR with Spatially Modulated Co-AttentionPeng Gao, Minghang Zheng, Xiaogang Wang, Jifeng Dai 等ICCV 2021 · 被引用 392 次
- Rethinking Transformer-based Set Prediction for Object DetectionZhiqing Sun, Shengcao Cao, Yiming Yang, Kris KitaniICCV 2021 · 被引用 381 次
- PedHunter: Occlusion Robust Pedestrian Detector in Crowded ScenesCheng Chi, Shifeng Zhang, Junliang Xing, Zhen Lei 等AAAI 2020 · 被引用 118 次
相关 Paper
- Detection in Crowded Scenes: One Proposal, Multiple PredictionsXuangeng Chu, Anlin Zheng, Xiangyu Zhang, Jian SunCVPR 2020
- QueryPose: Sparse Multi-Person Pose Regression via Spatial-Aware Part-Level QueryYabo Xiao, Kai Su, Xiaojuan Wang, Dongdong Yu 等NeurIPS 2022 · 被引用 32 次
- Dense Distinct Query for End-to-End Object DetectionShilong Zhang, Xinjiang Wang, Jiaqi Wang, Jiangmiao Pang 等CVPR 2023
- Multi-Instance Pose Networks: Rethinking Top-Down Pose EstimationRawal Khirodkar, Visesh Chari, Amit Agrawal, Ambrish TyagiICCV 2021 · 被引用 80 次
- Relational Learning for Joint Head and Human DetectionCheng Chi, Shifeng Zhang, Junliang Xing, Zhen Lei 等AAAI 2020 · 被引用 61 次
