How Useful Is Self-Supervised Pretraining for Visual Tasks?
Alejandro Newell, Jia Deng
Abstract
Recent advances have spurred incredible progress in self-supervised pretraining for vision. We investigate what factors may play a role in the utility of these pretraining methods for practitioners. To do this, we evaluate various self-supervised algorithms across a comprehensive array of synthetic datasets and downstream tasks. We prepare a suite of synthetic data that enables an endless supply of annotated images as well as full control over dataset difficulty. Our experiments offer insights into how the utility of self-supervision changes as the number of available labels grows as well as how the utility changes as a function of the downstream task and the properties of the training data. We also find that linear evaluation does not correlate with finetuning performance. Code and data is available at github.com/princeton-vl/selfstudy.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 767e1061-daec-468d-b6f8-81eea47c0b5cCited by top-tier papers19
- Weakly Supervised Semantic Segmentation for Large-Scale Point CloudYachao Zhang, Zhonghao Li, Yuan Xie, Yanyun Qu et al.AAAI 2021 · 116 citations
- A Closer Look at Self-Supervised Lightweight Vision TransformersShaoru Wang, Jin Gao, Zeming Li, Xiaoqin Zhang et al.ICML 2023 · 61 citations
- Understanding Negative Samples in Instance Discriminative Self-supervised Representation LearningKento Nozawa, Issei SatoNeurIPS 2021 · 56 citations
- Scaled ReLU Matters for Training Vision TransformersPichao Wang, Xue Wang, Hao Luo, Jingkai Zhou et al.AAAI 2022 · 55 citations
- Overwriting Pretrained Bias with Finetuning DataAngelina Wang, Olga RussakovskyICCV 2023 · 50 citations
Builds on6
- Data-Efficient Image Recognition with Contrastive Predictive CodingOlivier J. HénaffICML 2020 · 1,553 citations
- Rethinking ImageNet Pre-TrainingKaiming He, Ross B. Girshick, Piotr DollárICCV 2019 · 1,188 citations
- Local Aggregation for Unsupervised Learning of Visual EmbeddingsChengxu Zhuang, Alex Lin Zhai, Daniel YaminsICCV 2019 · 462 citations
- Scaling and Benchmarking Self-Supervised Visual Representation LearningPriya Goyal, Dhruv Mahajan, Abhinav Gupta, Ishan MisraICCV 2019 · 429 citations
- Momentum Contrast for Unsupervised Visual Representation LearningKaiming He, Haoqi Fan, Yuxin Wu, Saining Xie et al.CVPR 2020
Related papers
- Contrasting Contrastive Self-Supervised Representation Learning PipelinesKlemen Kotar, Gabriel Ilharco, Ludwig Schmidt, Kiana Ehsani et al.ICCV 2021 · 48 citations
- When Does Contrastive Visual Representation Learning Work?Elijah Cole, Xuan Yang, Kimberly Wilber, Oisin Mac Aodha et al.CVPR 2022 · 98 citations
- Insights into Pre-training via Simpler Synthetic TasksYuhuai Wu, Felix Li, Percy LiangNeurIPS 2022 · 29 citations
- The effectiveness of MAE pre-pretraining for billion-scale pretrainingMannat Singh, Quentin Duval, Kalyan Vasudev Alwala, Haoqi Fan et al.ICCV 2023 · 91 citations
- On Data Scaling in Masked Image ModelingZhenda Xie, Zheng Zhang, Yue Cao, Yutong Lin et al.CVPR 2023
