A Random CNN Sees Objects: One Inductive Bias of CNN and Its Applications
Yun-Hao Cao, Jianxin Wu
Abstract
This paper starts by revealing a surprising finding: without any learning, a randomly initialized CNN can localize objects surprisingly well. That is, a CNN has an inductive bias to naturally focus on objects, named as Tobias ("The object is at sight") in this paper. This empirical inductive bias is further analyzed and successfully applied to self-supervised learning. A CNN is encouraged to learn representations that focus on the foreground object, by transforming every image into various versions with different backgrounds, where the foreground and background separation is guided by Tobias. Experimental results show that the proposed Tobias significantly improves downstream tasks, especially for object detection. This paper also shows that Tobias has consistent improvements on training sets of different sizes, and is more resilient to changes in image augmentations. Code is available at https://github.com/CupidJay/Tobias .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 1b91cff9-c1aa-4d32-9fea-603413390c31Cited by top-tier papers4
- Multi-Label Self-Supervised Learning with Scene ImagesKe Zhu, Minghao Fu, Jianxin WuICCV 2023 · 21 citations
- Multiple Thinking Achieving Meta-Ability Decoupling for Object NavigationRonghao Dang, Lu Chen, Liuyi Wang, Zongtao He et al.ICML 2023 · 16 citations
- Activation Subspaces for Out-of-Distribution DetectionBaris Zöngür, Robin Hesse, Stefan RothICCV 2025 · 2 citations
- DisWOT: Student Architecture Search for Distillation WithOut TrainingPeijie Dong, Lujun Li, Zimian WeiCVPR 2023
Builds on10
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 24,064 citations
- Supervised Contrastive LearningPrannay Khosla, Piotr Teterwak, Chen Wang, Aaron Sarna et al.NeurIPS 2020 · 7,049 citations
- CutMix: Regularization Strategy to Train Strong Classifiers With Localizable FeaturesSangdoo Yun, Dongyoon Han, Sanghyuk Chun, Seong Joon Oh et al.ICCV 2019 · 5,843 citations
- Proving the Lottery Ticket Hypothesis: Pruning is All You NeedEran Malach, Gilad Yehudai, Shai Shalev-Shwartz, Ohad ShamirICML 2020 · 327 citations
- Un-mix: Rethinking Image Mixtures for Unsupervised Visual Representation LearningZhiqiang Shen, Zechun Liu, Zhuang Liu, Marios Savvides et al.AAAI 2022 · 117 citations
Related papers
- Instance Localization for Self-Supervised Detection PretrainingCeyuan Yang, Zhirong Wu, Bolei Zhou, Stephen LinCVPR 2021
- Distilling Localization for Self-Supervised Representation LearningNanxuan Zhao, Zhirong Wu, Rynson W. H. Lau, Stephen LinAAAI 2021 · 59 citations
- MOST: Multiple Object localization with Self-supervised Transformers for object discoverySai Saketh Rambhatla, Ishan Misra, Rama Chellappa, Abhinav ShrivastavaICCV 2023 · 21 citations
- Object-aware Contrastive Learning for Debiased Scene RepresentationSangwoo Mo, Hyunwoo Kang, Kihyuk Sohn, Chun-Liang Li et al.NeurIPS 2021 · 57 citations
- Spatially Consistent Representation LearningByungseok Roh, Wuhyun Shin, Ildoo Kim, Sungwoong KimCVPR 2021
