Robust Object Detection Under Occlusion With Context-Aware CompositionalNets
Angtian Wang, Yihong Sun, Adam Kortylewski, Alan L. Yuille
Abstract
Detecting partially occluded objects is a difficult task. Our experimental results show that deep learning approaches, such as Faster R-CNN, are not robust at object detection under occlusion. Compositional convolutional neural networks (CompositionalNets) have been shown to be robust at classifying occluded objects by explicitly representing the object as a composition of parts. In this work, we propose to overcome two limitations of Compositional-Nets which will enable them to detect partially occluded objects: 1) CompositionalNets, as well as other DCNN architectures, do not explicitly separate the representation of the context from the object itself. Under strong object occlusion, the influence of the context is amplified which can have severe negative effects for detection at test time. In order to overcome this, we propose to segment the context during training via bounding box annotations. We then use the segmentation to learn a context-aware CompositionalNet that disentangles the representation of the context and the object. 2) We extend the part-based voting scheme in Compo-sitionalNets to vote for the corners of the object's bounding box, which enables the model to reliably estimate bounding boxes for partially occluded objects. Our extensive experiments show that our proposed model can detect objects robustly, increasing the detection performance of strongly occluded vehicles from PASCAL3D+ and MS-COCO by 41% and 35% respectively in absolute performance relative to Faster R-CNN.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 56a93bbe-86fc-42e5-8293-d3ee2a230e9fCited by top-tier papers30
- NeMo: Neural Mesh Models of Contrastive Features for Robust 3D Pose EstimationAngtian Wang, Adam Kortylewski, Alan L. YuilleICLR 2021 · 53 citations
- UMC: A Unified Bandwidth-efficient and Multi-resolution based Collaborative Perception FrameworkTianhang Wang, Guang Chen, Kai Chen, Zhengfa Liu et al.ICCV 2023 · 46 citations
- Occlusion-Aware Video Object InpaintingLei Ke, Yu-Wing Tai, Chi-Keung TangICCV 2021 · 44 citations
- Learning Environment-Aware Affordance for 3D Articulated Object Manipulation under OcclusionsRuihai Wu, Kai Cheng, Yan Zhao, Chuanruo Ning et al.NeurIPS 2023 · 43 citations
- SwapMix: Diagnosing and Regularizing the Over-Reliance on Visual Context in Visual Question AnsweringVipul Gupta, Zhuowan Li, Adam Kortylewski, Chenyu Zhang et al.CVPR 2022 · 41 citations
Builds on2
- CutMix: Regularization Strategy to Train Strong Classifiers With Localizable FeaturesSangdoo Yun, Dongyoon Han, Sanghyuk Chun, Seong Joon Oh et al.ICCV 2019 · 5,843 citations
- Compositional Convolutional Neural Networks: A Deep Architecture With Innate Robustness to Partial OcclusionAdam Kortylewski, Ju He, Qing Liu, Alan L. YuilleCVPR 2020
Related papers
- Robust Instance Segmentation Through Reasoning About Multi-Object OcclusionXiaoding Yuan, Adam Kortylewski, Yihong Sun, Alan L. YuilleCVPR 2021
- Scaling 3D Compositional Models for Robust Classification and Pose EstimationXiaoding Yuan, Guofeng Zhang, Prakhar Kaushik, Artur Jesslen et al.ICCV 2025
- Deep Occlusion-Aware Instance Segmentation With Overlapping BiLayersLei Ke, Yu-Wing Tai, Chi-Keung TangCVPR 2021
- CBNet: A Novel Composite Backbone Network Architecture for Object DetectionYudong Liu, Yongtao Wang, Siwei Wang, Tingting Liang et al.AAAI 2020 · 266 citations
- Compositor: Bottom-Up Clustering and Compositing for Robust Part and Object SegmentationJu He, Jieneng Chen, Ming-Xian Lin, Qihang Yu et al.CVPR 2023
