Unsupervised Part Discovery from Contrastive Reconstruction
Subhabrata Choudhury, Iro Laina, Christian Rupprecht, Andrea Vedaldi
摘要
The goal of self-supervised visual representation learning is to learn strong, transferable image representations, with the majority of research focusing on object or scene level. On the other hand, representation learning at part level has received significantly less attention. In this paper, we propose an unsupervised approach to object part discovery and segmentation and make three contributions. First, we construct a proxy task through a set of objectives that encourages the model to learn a meaningful decomposition of the image into its parts. Secondly, prior work argues for reconstructing or clustering pre-computed features as a proxy to parts; we show empirically that this alone is unlikely to find meaningful parts; mainly because of their low resolution and the tendency of classification networks to spatially smear out information. We suggest that image reconstruction at the level of pixels can alleviate this problem, acting as a complementary cue. Lastly, we show that the standard evaluation based on keypoint regression does not correlate well with segmentation quality and thus introduce different metrics, NMI and ARI, that better characterize the decomposition of objects into parts. Our method yields semantic parts which are consistent across fine-grained but visually distinct categories, outperforming the state of the art on three benchmark datasets. Code is available at the project page: https://www.robots.ox.ac.uk/ vgg/research/unsup-parts/.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper37
- SemMAE: Semantic-Guided Masking for Learning Masked AutoencodersGang Li, Heliang Zheng, Daqing Liu, Chaoyue Wang 等NeurIPS 2022 · 被引用 188 次
- Deep Spectral Methods: A Surprisingly Strong Baseline for Unsupervised Semantic Segmentation and LocalizationLuke Melas-Kyriazi, Christian Rupprecht, Iro Laina, Andrea VedaldiCVPR 2022 · 被引用 132 次
- Self-Supervised Learning of Object Parts for Semantic SegmentationAdrian Ziegler, Yuki M. AsanoCVPR 2022 · 被引用 87 次
- LASSIE: Learning Articulated Shapes from Sparse Image Ensemble via 3D Part DiscoveryChun-Han Yao, Wei-Chih Hung, Yuanzhen Li, Michael Rubinstein 等NeurIPS 2022 · 被引用 83 次
- FeatureNeRF: Learning Generalizable NeRFs by Distilling Foundation ModelsJianglong Ye, Naiyan Wang, Xiaolong WangICCV 2023 · 被引用 56 次
它引用的顶会 Paper25
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 被引用 24,064 次
- Emerging Properties in Self-Supervised Vision TransformersMathilde Caron, Hugo Touvron, Ishan Misra, Hervé Jégou 等ICCV 2021 · 被引用 8,921 次
- Unsupervised Learning of Visual Features by Contrasting Cluster AssignmentsMathilde Caron, Ishan Misra, Julien Mairal, Priya Goyal 等NeurIPS 2020 · 被引用 5,249 次
- Data-Efficient Image Recognition with Contrastive Predictive CodingOlivier J. HénaffICML 2020 · 被引用 1,553 次
- Object-Centric Learning with Slot AttentionFrancesco Locatello, Dirk Weissenborn, Thomas Unterthiner, Aravindh Mahendran 等NeurIPS 2020 · 被引用 1,275 次
相关 Paper
- Unsupervised Part Segmentation Through Disentangling Appearance and ShapeShilong Liu, Lei Zhang, Xiao Yang, Hang Su 等CVPR 2021
- Unsupervised Semantic Segmentation with Self-supervised Object-centric RepresentationsAndrii Zadaianchuk, Matthäus Kleindessner, Yi Zhu, Francesco Locatello 等ICLR 2023 · 被引用 16 次
- Unsupervised Co-part Segmentation through AssemblyQingzhe Gao, Bin Wang, Libin Liu, Baoquan ChenICML 2021 · 被引用 16 次
- Hierarchy-Agnostic Unsupervised Segmentation: Parsing Semantic Image StructureSimone Rossetti, Fiora PirriNeurIPS 2024 · 被引用 2 次
- Towards Open-World Segmentation of PartsTai-Yu Pan, Qing Liu, Wei-Lun Chao, Brian L. PriceCVPR 2023
