Unsupervised Learning of Dense Visual Representations
Pedro O. Pinheiro, Amjad Almahairi, Ryan Y. Benmalek, Florian Golemo, Aaron C. Courville
Abstract
Contrastive self-supervised learning has emerged as a promising approach to unsupervised visual representation learning. In general, these methods learn global (image-level) representations that are invariant to different views (i.e., compositions of data augmentation) of the same image. However, many visual understanding tasks require dense (pixel-level) representations. In this paper, we propose View-Agnostic Dense Representation (VADeR) for unsupervised learning of dense representations. VADeR learns pixelwise representations by forcing local features to remain constant over different viewing conditions. Specifically, this is achieved through pixel-level contrastive learning: matching features (that is, features that describes the same location of the scene on different views) should be close in an embedding space, while non-matching features should be apart. VADeR provides a natural representation for dense prediction tasks and transfers well to downstream tasks. Our method outperforms ImageNet supervised pretraining (and strong unsupervised baselines) in multiple dense prediction tasks.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers76
- TS2Vec: Towards Universal Representation of Time SeriesZhihan Yue, Yujing Wang, Juanyong Duan, Tianmeng Yang et al.AAAI 2022 · 938 citations
- CRIS: CLIP-Driven Referring Image SegmentationZhaoqing Wang, Yu Lu, Qiang Li, Xunqiang Tao et al.CVPR 2022 · 337 citations
- Unsupervised Semantic Segmentation by Distilling Feature CorrespondencesMark Hamilton, Zhoutong Zhang, Bharath Hariharan, Noah Snavely et al.ICLR 2022 · 317 citations
- Pixel Contrastive-Consistent Semi-Supervised Semantic SegmentationYuanyi Zhong, Bodi Yuan, Hong Wu, Zhiqiang Yuan et al.ICCV 2021 · 210 citations
- Contrastive Learning for Label Efficient Semantic SegmentationXiangyun Zhao, Raviteja Vemulapalli, Philip Andrew Mansfield, Boqing Gong et al.ICCV 2021 · 200 citations
Builds on7
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 24,064 citations
- Data-Efficient Image Recognition with Contrastive Predictive CodingOlivier J. HénaffICML 2020 · 1,553 citations
- Local Aggregation for Unsupervised Learning of Visual EmbeddingsChengxu Zhuang, Alex Lin Zhai, Daniel YaminsICCV 2019 · 462 citations
- Scaling and Benchmarking Self-Supervised Visual Representation LearningPriya Goyal, Dhruv Mahajan, Abhinav Gupta, Ishan MisraICCV 2019 · 429 citations
- Unsupervised Learning of Landmarks by Descriptor Vector ExchangeJames Thewlis, Samuel Albanie, Hakan Bilen, Andrea VedaldiICCV 2019 · 70 citations
Related papers
- Propagate Yourself: Exploring Pixel-Level Consistency for Unsupervised Visual Representation LearningZhenda Xie, Yutong Lin, Zheng Zhang, Yue Cao et al.CVPR 2021
- Dense Contrastive Learning for Self-Supervised Visual Pre-TrainingXinlong Wang, Rufeng Zhang, Chunhua Shen, Tao Kong et al.CVPR 2021
- Dense Semantic Contrast for Self-Supervised Visual Representation LearningXiaoni Li, Yu Zhou, Yifei Zhang, Aoting Zhang et al.ACM MM 2021 · 35 citations
- Exploring Set Similarity for Dense Self-supervised Representation LearningZhaoqing Wang, Qiang Li, Guoxin Zhang, Pengfei Wan et al.CVPR 2022 · 33 citations
- Dense Contrastive Visual-Linguistic PretrainingLei Shi, Kai Shuang, Shijie Geng, Peng Gao et al.ACM MM 2021 · 12 citations
