PartDistillation: Learning Parts from Instance Segmentation
Jang Hyun Cho, Philipp Krähenbühl, Vignesh Ramanathan
Abstract
We present a scalable framework to learn part segmentation from object instance labels. State-of-the-art instance segmentation models contain a surprising amount of part information. However, much of this information is hidden from plain view. For each object instance, the part information is noisy, inconsistent, and incomplete. PartDistillation transfers the part information of an instance segmentation model into a part segmentation model through self-supervised self-training on a large dataset. The resulting segmentation model is robust, accurate, and generalizes well. We evaluate the model on various part segmentation datasets. Our model outperforms supervised part segmentation in zero-shot generalization performance by a large margin. Our model outperforms when finetuned on target datasets compared to supervised counterpart and other baselines especially in few-shot regime. Finally, our model provides a wider coverage of rare parts when evaluated over 10K object classes. Code is at https://github.com/facebookresearch/PartDistillation.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers3
- S2-Boost: Synergistic Semantic Boosting for Coarse-to-Fine Ensemble LearningGuanxiong He, Zheng Wang, Jie Wang, Liaoyuan Tang et al.AAAI 2026
- CALICO: Part-Focused Semantic Co-Segmentation with Large Vision-Language ModelsKiet A. Nguyen, Adheesh Sunil Juvekar, Tianjiao Yu, Muntasir Wahed et al.CVPR 2025
- Language-Conditioned Detection TransformerJang Hyun Cho, Philipp KrähenbühlCVPR 2024
Builds on18
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu et al.ICCV 2021 · 31,683 citations
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 24,064 citations
- Emerging Properties in Self-Supervised Vision TransformersMathilde Caron, Hugo Touvron, Ishan Misra, Hervé Jégou et al.ICCV 2021 · 8,921 citations
- Deformable DETR: Deformable Transformers for End-to-End Object DetectionXizhou Zhu, Weijie Su, Lewei Lu, Bin Li et al.ICLR 2021 · 7,353 citations
- A ConvNet for the 2020sZhuang Liu, Hanzi Mao, Chao-Yuan Wu, Christoph Feichtenhofer et al.CVPR 2022 · 6,782 citations
Related papers
- Towards Open-World Segmentation of PartsTai-Yu Pan, Qing Liu, Wei-Lun Chao, Brian L. PriceCVPR 2023
- MaskBooster: End-to-End Self-Training for Sparsely Supervised Instance SegmentationShida Zheng, Chenshu Chen, Xi Yang, Wenming TanAAAI 2023 · 1 citation
- FAPIS: A Few-Shot Anchor-Free Part-Based Instance SegmenterKhoi Nguyen, Sinisa TodorovicCVPR 2021
- Self-supervised Label Augmentation via Input TransformationsHankook Lee, Sung Ju Hwang, Jinwoo ShinICML 2020 · 218 citations
- PartDistill: 3D Shape Part Segmentation by Vision-Language Model DistillationArdian Umam, Cheng-Kun Yang, Min-Hung Chen, Jen-Hui Chuang et al.CVPR 2024 · 14 citations
