End-to-End Object Detection With Fully Convolutional Network
Jianfeng Wang, Lin Song, Zeming Li, Hongbin Sun, Jian Sun, Nanning Zheng
Abstract
Mainstream object detectors based on the fully convolutional network has achieved impressive performance. While most of them still need a hand-designed non-maximum suppression (NMS) post-processing, which impedes fully endto-end training. In this paper, we give the analysis of discarding NMS, where the results reveal that a proper label assignment plays a crucial role. To this end, for fully convolutional detectors, we introduce a Prediction-aware One-To-One (POTO) label assignment for classification to enable end-to-end detection, which obtains comparable performance with NMS. Besides, a simple 3D Max Filtering (3DMF) is proposed to utilize the multi-scale features and improve the discriminability of convolutions in the local region. With these techniques, our end-to-end framework achieves competitive performance against many state-ofthe-art detectors with NMS on COCO and CrowdHuman datasets. The code is available at https://github . com/Megvii-BaseDetection/DeFCN . * Equal contribution. † This work was done at Megvii Technology. * We remove its centerness branch to achieve a head-to-head comparison.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 9358209d-5c16-42ba-8e09-ce3ace68aa55Cited by top-tier papers37
- YOLOv10: Real-Time End-to-End Object DetectionAo Wang, Hui Chen, Lihao Liu, Kai Chen et al.NeurIPS 2024 · 6,113 citations
- YOLOv12: Attention-Centric Real-Time Object DetectorsYunjie Tian, Qixiang Ye, David S. DoermannNeurIPS 2025 · 2,652 citations
- Anchor DETR: Query Design for Transformer-Based DetectorYingming Wang, Xiangyu Zhang, Tong Yang, Jian SunAAAI 2022 · 567 citations
- GPT4Tools: Teaching Large Language Model to Use Tools via Self-instructionRui Yang, Lin Song, Yanwei Li, Sijie Zhao et al.NeurIPS 2023 · 340 citations
- Instances as QueriesYuxin Fang, Shusheng Yang, Xinggang Wang, Yu Li et al.ICCV 2021 · 331 citations
Builds on8
- FCOS: Fully Convolutional One-Stage Object DetectionZhi Tian, Chunhua Shen, Hao Chen, Tong HeICCV 2019 · 6,042 citations
- CenterNet: Keypoint Triplets for Object DetectionKaiwen Duan, Song Bai, Lingxi Xie, Honggang Qi et al.ICCV 2019 · 3,348 citations
- Fine-Grained Dynamic Head for Object DetectionLin Song, Yanwei Li, Zhengkai Jiang, Zeming Li et al.NeurIPS 2020 · 54 citations
- Rethinking Learnable Tree Filter for Generic Feature TransformLin Song, Yanwei Li, Zhengkai Jiang, Zeming Li et al.NeurIPS 2020 · 18 citations
- Learning From Noisy Anchors for One-Stage Object DetectionHengduo Li, Zuxuan Wu, Chen Zhu, Caiming Xiong et al.CVPR 2020
Related papers
- One-to-Few Label Assignment for End-to-End Dense DetectionShuai Li, Minghan Li, Ruihuang Li, Chenhang He et al.CVPR 2023
- What Makes for End-to-End Object Detection?Peize Sun, Yi Jiang, Enze Xie, Wenqi Shao et al.ICML 2021 · 106 citations
- Dense Distinct Query for End-to-End Object DetectionShilong Zhang, Xinjiang Wang, Jiaqi Wang, Jiangmiao Pang et al.CVPR 2023
- SOIT: Segmenting Objects with Instance-Aware TransformersXiaodong Yu, Dahu Shi, Xing Wei, Ye Ren et al.AAAI 2022 · 32 citations
- OTA: Optimal Transport Assignment for Object DetectionZheng Ge, Songtao Liu, Zeming Li, Osamu Yoshie et al.CVPR 2021
