Disentangle Your Dense Object Detector
Zehui Chen, Chenhongyi Yang, Qiaofei Li, Feng Zhao, Zheng-Jun Zha, Feng Wu
摘要
Deep learning-based dense object detectors have achieved great success in the past few years and have been applied to numerous multimedia applications such as video understanding. However, the current training pipeline for dense detectors is compromised to lots of conjunctions that may not hold. In this paper, we investigate three such important conjunctions: 1) only samples assigned as positive in classification head are used to train the regression head; 2) classification and regression share the same input feature and computational fields defined by the parallel head architecture; and 3) samples distributed in different feature pyramid layers are treated equally when computing the loss. We first carry out a series of pilot experiments to show disentangling such conjunctions can lead to persistent performance improvement. Then, based on these findings, we propose Disentangled Dense Object Detector (DDOD), in which simple and effective disentanglement mechanisms are designed and integrated into the current state-of-the-art dense object detectors. Extensive experiments on MS COCO benchmark show that our approach can lead to 2.0 mAP, 2.4 mAP and 2.2 mAP absolute improvements on RetinaNet, FCOS, and ATSS baselines with negligible extra overhead. Notably, our best model reaches 55.0 mAP on the COCO test-dev set and 93.5 AP on the hard subset of WIDER FACE, achieving new state-of-the-art performance on these two competitive benchmarks. Code is available at https://github.com/zehuichen123/DDOD.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper13
- Cascade-DETR: Delving into High-Quality Universal Object DetectionMingqiao Ye, Lei Ke, Siyuan Li, Yu-Wing Tai 等ICCV 2023 · 被引用 62 次
- DETRDistill: A Universal Knowledge Distillation Framework for DETR-familiesJiahao Chang, Shuo Wang, Hai-Ming Xu, Zehui Chen 等ICCV 2023 · 被引用 53 次
- Plug and Play Active Learning for Object DetectionChenhongyi Yang, Lichao Huang, Elliot J. CrowleyCVPR 2024 · 被引用 29 次
- Affine-Consistent Transformer for Multi-Class Cell Nuclei DetectionJunjia Huang, Haofeng Li, Xiang Wan, Guanbin LiICCV 2023 · 被引用 20 次
- STXD: Structural and Temporal Cross-Modal Distillation for Multi-View 3D Object DetectionSujin Jang, Dae Ung Jo, Sung Ju Hwang, Dongwook Lee 等NeurIPS 2023 · 被引用 19 次
它引用的顶会 Paper21
- FCOS: Fully Convolutional One-Stage Object DetectionZhi Tian, Chunhua Shen, Hao Chen, Tong HeICCV 2019 · 被引用 6,042 次
- Distance-IoU Loss: Faster and Better Learning for Bounding Box RegressionZhaohui Zheng, Ping Wang, Wei Liu, Jinze Li 等AAAI 2020 · 被引用 4,823 次
- CenterNet: Keypoint Triplets for Object DetectionKaiwen Duan, Song Bai, Lingxi Xie, Honggang Qi 等ICCV 2019 · 被引用 3,348 次
- Generalized Focal Loss: Learning Qualified and Distributed Bounding Boxes for Dense Object DetectionXiang Li, Wenhai Wang, Lijun Wu, Shuo Chen 等NeurIPS 2020 · 被引用 2,118 次
- RepPoints: Point Set Representation for Object DetectionZe Yang, Shaohui Liu, Han Hu, Liwei Wang 等ICCV 2019 · 被引用 1,056 次
相关 Paper
- Dense Distinct Query for End-to-End Object DetectionShilong Zhang, Xinjiang Wang, Jiaqi Wang, Jiangmiao Pang 等CVPR 2023
- VarifocalNet: An IoU-Aware Dense Object DetectorHaoyang Zhang, Ying Wang, Feras Dayoub, Niko SünderhaufCVPR 2021
- Sparse R-CNN: End-to-End Object Detection With Learnable ProposalsPeize Sun, Rufeng Zhang, Yi Jiang, Tao Kong 等CVPR 2021
- D2Det: Towards High Quality Object Detection and Instance SegmentationJiale Cao, Hisham Cholakkal, Rao Muhammad Anwer, Fahad Shahbaz Khan 等CVPR 2020
- DetCo: Unsupervised Contrastive Learning for Object DetectionEnze Xie, Jian Ding, Wenhai Wang, Xiaohang Zhan 等ICCV 2021 · 被引用 364 次
