Rethinking Classification and Localization for Object Detection
Yue Wu, Yinpeng Chen, Lu Yuan, Zicheng Liu, Lijuan Wang, Hongzhi Li, Yun Fu
摘要
Two head structures (i.e. fully connected head and convolution head) have been widely used in R-CNN based detectors for classification and localization tasks. However, there is a lack of understanding of how does these two head structures work for these two tasks. To address this issue, we perform a thorough analysis and find an interesting fact that the two head structures have opposite preferences towards the two tasks. Specifically, the fully connected head (fc-head) is more suitable for the classification task, while the convolution head (conv-head) is more suitable for the localization task. Furthermore, we examine the output feature maps of both heads and find that fc-head has more spatial sensitivity than conv-head. Thus, fc-head has more capability to distinguish a complete object from part of an object, but is not robust to regress the whole object. Based upon these findings, we propose a Double-Head method, which has a fully connected head focusing on classification and a convolution head for bounding box regression. Without bells and whistles, our method gains +3.5 and +2.8 AP on MS COCO dataset from Feature Pyramid Network (FPN) baselines with ResNet-50 and ResNet-101 backbones, respectively.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper33
- DeFRCN: Decoupled Faster R-CNN for Few-Shot Object DetectionLimeng Qiao, Yuxuan Zhao, Zhiyuan Li, Xi Qiu 等ICCV 2021 · 被引用 298 次
- Disentangle Your Dense Object DetectorZehui Chen, Chenhongyi Yang, Qiaofei Li, Feng Zhao 等ACM MM 2021 · 被引用 189 次
- SARDet-100K: Towards Open-Source Benchmark and ToolKit for Large-Scale SAR Object DetectionYuxuan Li, Xiang Li, Weijie Li, Qibin Hou 等NeurIPS 2024 · 被引用 145 次
- Real-time Object Detection for Streaming PerceptionJinrong Yang, Songtao Liu, Zeming Li, Xiaoping Li 等CVPR 2022 · 被引用 61 次
- ESTextSpotter: Towards Better Scene Text Spotting with Explicit Synergy in TransformerMingxin Huang, Jiaxin Zhang, Dezhi Peng, Hao Lu 等ICCV 2023 · 被引用 44 次
它引用的顶会 Paper4
- CenterNet: Keypoint Triplets for Object DetectionKaiwen Duan, Song Bai, Lingxi Xie, Honggang Qi 等ICCV 2019 · 被引用 3,348 次
- Scale-Aware Trident Networks for Object DetectionYanghao Li, Yuntao Chen, Naiyan Wang, Zhaoxiang ZhangICCV 2019 · 被引用 1,031 次
- Towards Adversarially Robust Object DetectionHaichao Zhang, Jianyu WangICCV 2019 · 被引用 152 次
- Learning to Rank Proposals for Object DetectionZhiyu Tan, Xuecheng Nie, Qi Qian, Nan Li 等ICCV 2019 · 被引用 53 次
相关 Paper
- AFD-Net: Adaptive Fully-Dual Network for Few-Shot Object DetectionLongyao Liu, Bo Ma, Yulin Zhang, Xin Yi 等ACM MM 2021 · 被引用 19 次
- D2Det: Towards High Quality Object Detection and Instance SegmentationJiale Cao, Hisham Cholakkal, Rao Muhammad Anwer, Fahad Shahbaz Khan 等CVPR 2020
- Auto-FPN: Automatic Network Architecture Adaptation for Object Detection Beyond ClassificationHang Xu, Lewei Yao, Zhenguo Li, Xiaodan Liang 等ICCV 2019 · 被引用 197 次
- RCNet: Reverse Feature Pyramid and Cross-scale Shift Network for Object DetectionZhuofan Zong, Qianggang Cao, Biao LengACM MM 2021 · 被引用 22 次
- Dynamic Head: Unifying Object Detection Heads With AttentionsXiyang Dai, Yinpeng Chen, Bin Xiao, Dongdong Chen 等CVPR 2021
