Anchor-Intermediate Detector: Decoupling and Coupling Bounding Boxes for Accurate Object Detection
Yilong Lv, Min Li, Yujie He, Zhuzhen He, Shaopeng Li, Aitao Yang
Abstract
Anchor-based detectors have been continuously developed for object detection. However, the individual anchor box makes it difficult to predict the boundary's offset accurately. Instead of taking each bounding box as a closed individual, we consider using multiple boxes together to get prediction boxes. To this end, this paper proposes the Box Decouple-Couple(BDC) strategy in the inference, which no longer discards the overlapping boxes, but decouples the corner points of these boxes. Then, according to each corner's score, we couple the corner points to select the most accurate corner pairs. To meet the BDC strategy, a simple but novel model is designed named the Anchor-Intermediate Detector(AID), which contains two head networks, i.e., an anchor-based head and an anchorfree Corner-aware head. The corner-aware head is able to score the corners of each bounding box to facilitate the coupling between corner points. Extensive experiments on MS COCO show that the proposed anchor-intermediate detector respectively outperforms their baseline RetinaNet and GFL method by ∼2.4 and ∼1.2 AP on the MS COCO testdev dataset without any bells and whistles. Code is available at: https://github.com/YilongLv/AID .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 8ff4aff4-b435-4e55-93cf-d193c77ae5c4Builds on12
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu et al.ICCV 2021 · 31,683 citations
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- FCOS: Fully Convolutional One-Stage Object DetectionZhi Tian, Chunhua Shen, Hao Chen, Tong HeICCV 2019 · 6,042 citations
- Pyramid Vision Transformer: A Versatile Backbone for Dense Prediction without ConvolutionsWenhai Wang, Enze Xie, Xiang Li, Deng-Ping Fan et al.ICCV 2021 · 4,909 citations
- CenterNet: Keypoint Triplets for Object DetectionKaiwen Duan, Song Bai, Lingxi Xie, Honggang Qi et al.ICCV 2019 · 3,348 citations
Related papers
- CrossDet: Crossline Representation for Object DetectionHeqian Qiu, Hongliang Li, Qingbo Wu, Jianhua Cui et al.ICCV 2021 · 17 citations
- Anchor DETR: Query Design for Transformer-Based DetectorYingming Wang, Xiangyu Zhang, Tong Yang, Jian SunAAAI 2022 · 567 citations
- Gradient Corner Pooling for Keypoint-Based Object DetectionXuyang Li, Xuemei Xie, Mingxuan Yu, Jiakai Luo et al.AAAI 2023 · 1 citation
- HAMBox: Delving Into Mining High-Quality Anchors on Face DetectionYang Liu, Xu Tang, Junyu Han, Jingtuo Liu et al.CVPR 2020
- CentripetalNet: Pursuing High-Quality Keypoint Pairs for Object DetectionZhiwei Dong, Guoxuan Li, Yue Liao, Fei Wang et al.CVPR 2020
