Self-Supervised Visual Representation Learning from Hierarchical Grouping
Xiao Zhang, Michael Maire
摘要
We create a framework for bootstrapping visual representation learning from a primitive visual grouping capability. We operationalize grouping via a contour detector that partitions an image into regions, followed by merging of those regions into a tree hierarchy. A small supervised dataset suffices for training this grouping primitive. Across a large unlabeled dataset, we apply this learned primitive to automatically predict hierarchical region structure. These predictions serve as guidance for self-supervised contrastive feature learning: we task a deep network with producing per-pixel embeddings whose pairwise distances respect the region hierarchy. Experiments demonstrate that our approach can serve as state-of-the-art generic pre-training, benefiting downstream tasks. We additionally explore applications to semantic region search and video-based object instance tracking.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper35
- Unsupervised Semantic Segmentation by Contrasting Object Mask ProposalsWouter Van Gansbeke, Simon Vandenhende, Stamatios Georgoulis, Luc Van GoolICCV 2021 · 被引用 285 次
- MST: Masked Self-Supervised Transformer for Visual RepresentationZhaowen Li, Zhiyang Chen, Fan Yang, Wei Li 等NeurIPS 2021 · 被引用 194 次
- Efficient Visual Pretraining with Contrastive DetectionOlivier J. Hénaff, Skanda Koppula, Jean-Baptiste Alayrac, Aäron van den Oord 等ICCV 2021 · 被引用 186 次
- ReCo: Retrieve and Co-segment for Zero-shot TransferGyungin Shin, Weidi Xie, Samuel AlbanieNeurIPS 2022 · 被引用 160 次
- Region-aware Contrastive Learning for Semantic SegmentationHanzhe Hu, Jinshi Cui, Liwei WangICCV 2021 · 被引用 132 次
它引用的顶会 Paper4
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 被引用 24,064 次
- Local Aggregation for Unsupervised Learning of Visual EmbeddingsChengxu Zhuang, Alex Lin Zhai, Daniel YaminsICCV 2019 · 被引用 462 次
- SegSort: Segmentation by Discriminative Sorting of SegmentsJyh-Jing Hwang, Stella X. Yu, Jianbo Shi, Maxwell D. Collins 等ICCV 2019 · 被引用 160 次
- Momentum Contrast for Unsupervised Visual Representation LearningKaiming He, Haoqi Fan, Yuxin Wu, Saining Xie 等CVPR 2020
相关 Paper
- Self-Supervised Visual Representation Learning with Semantic GroupingXin Wen, Bingchen Zhao, Anlin Zheng, Xiangyu Zhang 等NeurIPS 2022 · 被引用 104 次
- Dense Semantic Contrast for Self-Supervised Visual Representation LearningXiaoni Li, Yu Zhou, Yifei Zhang, Aoting Zhang 等ACM MM 2021 · 被引用 35 次
- Point-Level Region Contrast for Object Detection Pre-TrainingYutong Bai, Xinlei Chen, Alexander Kirillov, Alan L. Yuille 等CVPR 2022 · 被引用 43 次
- Train a One-Million-Way Instance Classifier for Unsupervised Visual Representation LearningYu Liu, Lianghua Huang, Pan Pan, Bin Wang 等AAAI 2021 · 被引用 3 次
- Object Concepts Emerge from MotionHaoqian Liang, Xiaohui Wang, Zhichao Li, Ya Yang 等NeurIPS 2025
