POD: Practical Object Detection With Scale-Sensitive Network
Junran Peng, Ming Sun, Zhaoxiang Zhang, Tieniu Tan, Junjie Yan
Abstract
Scale-sensitive object detection remains a challenging task, where most of the existing methods could not learn it explicitly and are not robust to scale variance. In addition, the most existing methods are less efficient during training or slow during inference, which are not friendly to real-time applications. In this paper, we propose a practical object detection method with scale-sensitive network. Our method first predicts a global continuous scale , which is shared by all position, for each convolution filter of each network stage. To effectively learn the scale, we average the spatial features and distill the scale from channels. For fast-deployment, we propose a scale decomposition method that transfers the robust fractional scale into combination of fixed integral scales for each convolution filter, which exploits the dilated convolution. We demonstrate it on one-stage and two-stage algorithms under different configurations. For practical applications, training of our method is of efficiency and simplicity which gets rid of complex data sampling or optimize strategy. During testing, the proposed method requires no extra operation and is very supportive of hardware acceleration like TensorRT and TVM. On the COCO test-dev, our model could achieve a 41.5 mAP on one-stage detector and 42.1 mAP on twostage detectors based on ResNet-101, outperforming baselines by 2.4 and 2.1 respectively without extra FLOPS.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 93a080fe-df4c-45e8-af2e-b043ad72320dCited by top-tier papers7
- Robust Small-scale Pedestrian Detection with Cued Recall via Memory LearningJung Uk Kim, Sungjune Park, Yong Man RoICCV 2021 · 61 citations
- Log-Polar Space Convolution LayersBing Su, Ji-Rong WenNeurIPS 2022 · 3 citations
- Large-Scale Object Detection in the Wild From Imbalanced Multi-LabelsJunran Peng, Xingyuan Bu, Ming Sun, Zhaoxiang Zhang et al.CVPR 2020
- Multiple Anchor Learning for Visual Object DetectionWei Ke, Tianliang Zhang, Zeyi Huang, Qixiang Ye et al.CVPR 2020
- Learning From Noisy Anchors for One-Stage Object DetectionHengduo Li, Zuxuan Wu, Chen Zhu, Caiming Xiong et al.CVPR 2020
Builds on1
Related papers
- Deformable Part Region Learning for Object DetectionSeung-Hwan BaeAAAI 2022 · 5 citations
- Scale-Aware Trident Networks for Object DetectionYanghao Li, Yuntao Chen, Naiyan Wang, Zhaoxiang ZhangICCV 2019 · 1,031 citations
- ScaleNet - Improve CNNs through Recursively Rescaling ObjectsXingyi Li, Zhongang Qi, Xiaoli Z. Fern, Fuxin LiAAAI 2020 · 1 citation
- YOLO-ULM: Ultra-Lightweight Models for Real-Time Object DetectionShasha Han, Chong Li, Xinning Wang, Xuebo LiCVPR 2026
- Focal and Global Knowledge Distillation for DetectorsZhendong Yang, Zhe Li, Xiaohu Jiang, Yuan Gong et al.CVPR 2022 · 325 citations
