Dense Semantic Contrast for Self-Supervised Visual Representation Learning
Xiaoni Li, Yu Zhou, Yifei Zhang, Aoting Zhang, Wei Wang, Ning Jiang, Haiying Wu, Weiping Wang
摘要
Self-supervised representation learning for visual pre-training has achieved remarkable success with sample (instance or pixel) discrimination and semantics discovery of instance, whereas there still exists a non-negligible gap between pre-trained model and downstream dense prediction tasks. Concretely, these downstream tasks require more accurate representation, in other words, the pixels from the same object must belong to a shared semantic category, which is lacking in the previous methods. In this work, we present Dense Semantic Contrast (DSC) for modeling semantic category decision boundaries at a dense level to meet the requirement of these tasks. Furthermore, we propose a dense cross-image semantic contrastive learning framework for multi-granularity representation learning. Specially, we explicitly explore the semantic structure of the dataset by mining relations among pixels from different perspectives. For intra-image relation modeling, we discover pixel neighbors from multiple views. And for inter-image relations, we enforce pixel representation from the same semantic class to be more similar than the representation from different classes in one mini-batch. Experimental results show that our DSC model outperforms state-of-the-art methods when transferring to downstream dense prediction tasks, including object detection, semantic segmentation, and instance segmentation. Code will be made available.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper8
- Mutual Contrastive Learning for Visual Representation LearningChuanguang Yang, Zhulin An, Linhang Cai, Yongjun XuAAAI 2022 · 被引用 95 次
- Self-Supervised Learning of Object Parts for Semantic SegmentationAdrian Ziegler, Yuki M. AsanoCVPR 2022 · 被引用 87 次
- Semantics-Consistent Feature Search for Self-Supervised Visual Representation LearningKaiyou Song, Shan Zhang, Zimeng Luo, Tong Wang 等ICCV 2023 · 被引用 10 次
- Representation Learning by Detecting Incorrect Location EmbeddingsSepehr Sameni, Simon Jenni, Paolo FavaroAAAI 2023 · 被引用 8 次
- Sound and Visual Representation Learning with Multiple Pretraining TasksArun Balajee Vasudevan, Dengxin Dai, Luc Van GoolCVPR 2022 · 被引用 4 次
它引用的顶会 Paper17
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 被引用 24,064 次
- Bootstrap Your Own Latent - A New Approach to Self-Supervised LearningJean-Bastien Grill, Florian Strub, Florent Altché, Corentin Tallec 等NeurIPS 2020 · 被引用 9,171 次
- Unsupervised Learning of Visual Features by Contrasting Cluster AssignmentsMathilde Caron, Ishan Misra, Julien Mairal, Priya Goyal 等NeurIPS 2020 · 被引用 5,249 次
- Invariant Information Clustering for Unsupervised Image Classification and SegmentationXu Ji, Andrea Vedaldi, João F. HenriquesICCV 2019 · 被引用 956 次
- Self-labelling via simultaneous clustering and representation learningYuki Markus Asano, Christian Rupprecht, Andrea VedaldiICLR 2020 · 被引用 873 次
相关 Paper
- Dense Contrastive Learning for Self-Supervised Visual Pre-TrainingXinlong Wang, Rufeng Zhang, Chunhua Shen, Tao Kong 等CVPR 2021
- Propagate Yourself: Exploring Pixel-Level Consistency for Unsupervised Visual Representation LearningZhenda Xie, Yutong Lin, Zheng Zhang, Yue Cao 等CVPR 2021
- Exploring Set Similarity for Dense Self-supervised Representation LearningZhaoqing Wang, Qiang Li, Guoxin Zhang, Pengfei Wan 等CVPR 2022 · 被引用 33 次
- DetCo: Unsupervised Contrastive Learning for Object DetectionEnze Xie, Jian Ding, Wenhai Wang, Xiaohang Zhan 等ICCV 2021 · 被引用 364 次
- DenseCLIP: Language-Guided Dense Prediction with Context-Aware PromptingYongming Rao, Wenliang Zhao, Guangyi Chen, Yansong Tang 等CVPR 2022 · 被引用 527 次
