Compositional Convolutional Neural Networks: A Deep Architecture With Innate Robustness to Partial Occlusion
Adam Kortylewski, Ju He, Qing Liu, Alan L. Yuille
摘要
Recent findings show that deep convolutional neural networks (DCNNs) do not generalize well under partial occlusion. Inspired by the success of compositional models at classifying partially occluded objects, we propose to integrate compositional models and DCNNs into a unified deep model with innate robustness to partial occlusion. We term this architecture Compositional Convolutional Neural Network. In particular, we propose to replace the fully connected classification head of a DCNN with a differentiable compositional model. The generative nature of the compositional model enables it to localize occluders and subsequently focus on the non-occluded parts of the object. We conduct classification experiments on artificially occluded images as well as real images of partially occluded objects from the MS-COCO dataset. The results show that DC-NNs do not classify occluded objects robustly, even when trained with data that is strongly augmented with partial occlusions. Our proposed model outperforms standard DC-NNs by a large margin at classifying partially occluded objects, even when it has not been exposed to occluded objects during training. Additional experiments demonstrate that CompositionalNets can also localize the occluders accurately, despite being trained with class labels only. The code used in this work is publicly available 1 .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper21
- NeMo: Neural Mesh Models of Contrastive Features for Robust 3D Pose EstimationAngtian Wang, Adam Kortylewski, Alan L. YuilleICLR 2021 · 被引用 53 次
- 3D-Aware Neural Body Fitting for Occlusion Robust 3D Human Pose EstimationYi Zhang, Pengliang Ji, Angtian Wang, Jieru Mei 等ICCV 2023 · 被引用 44 次
- Occlusion-Aware Video Object InpaintingLei Ke, Yu-Wing Tai, Chi-Keung TangICCV 2021 · 被引用 44 次
- SwapMix: Diagnosing and Regularizing the Over-Reliance on Visual Context in Visual Question AnsweringVipul Gupta, Zhuowan Li, Adam Kortylewski, Chenyu Zhang 等CVPR 2022 · 被引用 41 次
- Using latent space regression to analyze and leverage compositionality in GANsLucy Chai, Jonas Wulff, Phillip IsolaICLR 2021 · 被引用 30 次
它引用的顶会 Paper1
相关 Paper
- Robust Object Detection Under Occlusion With Context-Aware CompositionalNetsAngtian Wang, Yihong Sun, Adam Kortylewski, Alan L. YuilleCVPR 2020
- Robust Instance Segmentation Through Reasoning About Multi-Object OcclusionXiaoding Yuan, Adam Kortylewski, Yihong Sun, Alan L. YuilleCVPR 2021
- Deep Occlusion-Aware Instance Segmentation With Overlapping BiLayersLei Ke, Yu-Wing Tai, Chi-Keung TangCVPR 2021
- The surprising impact of mask-head architecture on novel class segmentationVighnesh Birodkar, Zhichao Lu, Siyang Li, Vivek Rathod 等ICCV 2021 · 被引用 32 次
- Scaling 3D Compositional Models for Robust Classification and Pose EstimationXiaoding Yuan, Guofeng Zhang, Prakhar Kaushik, Artur Jesslen 等ICCV 2025
