Deformable Part Region Learning for Object Detection
Seung-Hwan Bae
Abstract
In a convolutional object detector, the detection accuracy can be degraded often due to the low feature discriminability caused by geometric variation or transformation of an object. In this paper, we propose a deformable part region learning in order to allow decomposed part regions to be deformable according to geometric transformation of an object. To this end, we introduce trainable geometric parameters for the location of each part model. Because the ground truth of the part models is not available, we design classification and mask losses for part models, and learn the geometric parameters by minimizing an integral loss including those part losses. As a result, we can train a deformable part region network without extra super-vision, and make each part model deformable according to object scale variation. Furthermore, for improving cascade object detection and instance segmentation, we present a Cascade deformable part region architecture which can refine whole and part detections iteratively in the cascade manner. Without bells and whistles, our implementation of a Cascade deformable part region detector achieves better detection and segmentation mAPs on COCO and VOC datasets, compared to the recent cascade and other state-of-the-art detectors.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext c9996aa3-e635-4e1f-944f-3a65ca5d0114Cited by top-tier papers1
Ask how each one uses itBuilds on11
- FCOS: Fully Convolutional One-Stage Object DetectionZhi Tian, Chunhua Shen, Hao Chen, Tong HeICCV 2019 · 6,042 citations
- YOLACT: Real-Time Instance SegmentationDaniel Bolya, Chong Zhou, Fanyi Xiao, Yong Jae LeeICCV 2019 · 2,075 citations
- WSOD2: Learning Bottom-Up and Top-Down Objectness Distillation for Weakly-Supervised Object DetectionZhaoyang Zeng, Bei Liu, Jianlong Fu, Hongyang Chao et al.ICCV 2019 · 162 citations
- SCNet: Training Inference Sample Consistency for Instance SegmentationThang Vu, Haeyong Kang, Chang D. YooAAAI 2021 · 111 citations
- CentripetalNet: Pursuing High-Quality Keypoint Pairs for Object DetectionZhiwei Dong, Guoxuan Li, Yue Liao, Fei Wang et al.CVPR 2020
Related papers
- Towards Accurate Facial Landmark Detection via Cascaded TransformersHui Li, Zidong Guo, Seon-Min Rhee, Seungju Han et al.CVPR 2022 · 45 citations
- Construct Effective Geometry Aware Feature Pyramid Network for Multi-Scale Object DetectionJinpeng Dong, Yuhao Huang, Songyi Zhang, Shitao Chen et al.AAAI 2022 · 9 citations
- POD: Practical Object Detection With Scale-Sensitive NetworkJunran Peng, Ming Sun, Zhaoxiang Zhang, Tieniu Tan et al.ICCV 2019 · 23 citations
- DPT: Deformable Patch-based Transformer for Visual RecognitionZhiyang Chen, Yousong Zhu, Chaoyang Zhao, Guosheng Hu et al.ACM MM 2021 · 118 citations
- Sparse DETR: Efficient End-to-End Object Detection with Learnable SparsityByungseok Roh, Jaewoong Shin, Wuhyun Shin, Saehoon KimICLR 2022 · 256 citations
