Contrastive Attention Maps for Self-supervised Co-localization
Minsong Ki, Youngjung Uh, Junsuk Choe, Hyeran Byun
Abstract
The goal of unsupervised co-localization is to locate the object in a scene under the assumptions that 1) the dataset consists of only one superclass, e.g., birds, and 2) there are no human-annotated labels in the dataset. The most recent method achieves impressive co-localization performance by employing self-supervised representation learning approaches such as predicting rotation. In this paper, we introduce a new contrastive objective directly on the attention maps to enhance co-localization performance. Our contrastive loss function exploits rich information of location, which induces the model to activate the extent of the object effectively. In addition, we propose a pixel-wise attention pooling that selectively aggregates the feature map regarding their magnitudes across channels. Our methods are simple and shown effective by extensive qualitative and quantitative evaluation, achieving state-of-the-art co-localization performances by large margins on four datasets: CUB-200-2011, Stanford Cars, FGVC-Aircraft, and Stanford Dogs. Our code will be publicly available online for the research community.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext a62be64d-6b8e-4696-9c0a-c921cb41a442Cited by top-tier papers2
- Unsupervised Object Localization with Representer Point SelectionYeonghwan Song, Seokwoo Jang, Dina Katabi, Jeany SonICCV 2023 · 4 citations
- Contrastive Attention Networks for Attribution of Early Modern PrintNikolai Vogler, Kartik Goyal, Kishore PV Reddy, Elizaveta Pertseva et al.AAAI 2023 · 1 citation
Builds on10
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 24,064 citations
- Big Self-Supervised Models are Strong Semi-Supervised LearnersTing Chen, Simon Kornblith, Kevin Swersky, Mohammad Norouzi et al.NeurIPS 2020 · 2,611 citations
- Group-Wise Deep Object Co-Segmentation With Co-Attention Recurrent Neural NetworkBo Li, Zhengxing Sun, Qian Li, Yunjie Wu et al.ICCV 2019 · 62 citations
- Deep Object Co-Segmentation via Spatial-Semantic Network ModulationKaihua Zhang, Jin Chen, Bo Liu, Qingshan LiuAAAI 2020 · 43 citations
- PsyNet: Self-Supervised Approach to Object Localization Using Point Symmetric TransformationKyungjune Baek, Minhyun Lee, Hyunjung ShimAAAI 2020 · 37 citations
Related papers
- Self-Supervised Object Localization with Joint Graph PartitionYukun Su, Guosheng Lin, Yun Hao, Yiwen Cao et al.AAAI 2022 · 17 citations
- Towards End-to-End Unsupervised Saliency Detection with Self-Supervised Top-Down ContextYicheng Song, Shuyong Gao, Haozhe Xing, Yiting Cheng et al.ACM MM 2023 · 1 citation
- Point-Level Region Contrast for Object Detection Pre-TrainingYutong Bai, Xinlei Chen, Alexander Kirillov, Alan L. Yuille et al.CVPR 2022 · 43 citations
- Spatially Consistent Representation LearningByungseok Roh, Wuhyun Shin, Ildoo Kim, Sungwoong KimCVPR 2021
- DetCo: Unsupervised Contrastive Learning for Object DetectionEnze Xie, Jian Ding, Wenhai Wang, Xiaohang Zhan et al.ICCV 2021 · 364 citations
