Mitigating the Effect of Incidental Correlations on Part-based Learning
Gaurav Bhatt, Deepayan Das, Leonid Sigal, Vineeth N. Balasubramanian
Abstract
Intelligent systems possess a crucial characteristic of breaking complicated problems into smaller reusable components or parts and adjusting to new tasks using these part representations. However, current part-learners encounter difficulties in dealing with incidental correlations resulting from the limited observations of objects that may appear only in specific arrangements or with specific backgrounds. These incidental correlations may have a detrimental impact on the generalization and interpretability of learned part representations. This study asserts that part-based representations could be more interpretable and generalize better with limited data, employing two innovative regularization methods. The first regularization separates foreground and background information's generative process via a unique mixture-of-parts formulation. Structural constraints are imposed on the parts using a weakly-supervised loss, guaranteeing that the mixture-of-parts for foreground and background entails soft, object-agnostic masks. The second regularization assumes the form of a distillation loss, ensuring the invariance of the learned parts to the incidental background correlations. Furthermore, we incorporate sparse and orthogonal constraints to facilitate learning high-quality part representations. By reducing the impact of incidental background correlations on the learned parts, we exhibit state-of-the-art (SoTA) performance on few-shot learning tasks on benchmark datasets, including MiniImagenet, TieredImageNet, and FC100. We also demonstrate that the part-based representations acquired through our approach generalize better than existing techniques, even under domain shifts of the background and common data corruption on the ImageNet-9 dataset. The implementation is available on GitHub: https://github.com/GauravBh1010tt/DPViT.git
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext fa20c13f-15ae-4968-b52c-e602d0441b33Builds on21
- Emerging Properties in Self-Supervised Vision TransformersMathilde Caron, Hugo Touvron, Ishan Misra, Hervé Jégou et al.ICCV 2021 · 8,921 citations
- Noise or Signal: The Role of Image Backgrounds in Object RecognitionKai Yuanqing Xiao, Logan Engstrom, Andrew Ilyas, Aleksander MadryICLR 2021 · 451 citations
- Image BERT Pre-training with Online TokenizerJinghao Zhou, Chen Wei, Huiyu Wang, Wei Shen et al.ICLR 2022 · 287 citations
- Learning Compositional Representations for Few-Shot RecognitionPavel Tokmakov, Yu-Xiong Wang, Martial HebertICCV 2019 · 133 citations
- Matching Feature Sets for Few-Shot Image ClassificationArman Afrasiyabi, Hugo Larochelle, Jean-François Lalonde, Christian GagnéCVPR 2022 · 124 citations
Related papers
- Revisiting Pose-Normalization for Fine-Grained Few-Shot RecognitionLuming Tang, Davis Wertheimer, Bharath HariharanCVPR 2020
- Leveraging GAN Priors for Few-Shot Part SegmentationMengya Han, Heliang Zheng, Chaoyue Wang, Yong Luo et al.ACM MM 2022 · 5 citations
- Exploring Tuning Characteristics of Ventral Stream's Neurons for Few-Shot Image ClassificationLintao Dong, Wei Zhai, Zheng-Jun ZhaAAAI 2023 · 12 citations
- PartDistillation: Learning Parts from Instance SegmentationJang Hyun Cho, Philipp Krähenbühl, Vignesh RamanathanCVPR 2023
- Boosting Few-Shot Learning via Attentive Feature RegularizationXingyu Zhu, Shuo Wang, Jinda Lu, Yanbin Hao et al.AAAI 2024 · 30 citations
