LCTR: On Awakening the Local Continuity of Transformer for Weakly Supervised Object Localization
Zhiwei Chen, Changan Wang, Yabiao Wang, Guannan Jiang, Yunhang Shen, Ying Tai, Chengjie Wang, Wei Zhang, Liujuan Cao
Abstract
Weakly supervised object localization (WSOL) aims to learn object localizer solely by using image-level labels. The convolution neural network (CNN) based techniques often result in highlighting the most discriminative part of objects while ignoring the entire object extent. Recently, the transformer architecture has been deployed to WSOL to capture the longrange feature dependencies with self-attention mechanism and multilayer perceptron structure. Nevertheless, transformers lack the locality inductive bias inherent to CNNs and therefore may deteriorate local feature details in WSOL. In this paper, we propose a novel framework built upon the transformer, termed LCTR (Local Continuity TRansformer), which targets at enhancing the local perception capability of global features among long-range feature dependencies. To this end, we propose a relational patch-attention module (RPAM), which considers cross-patch information on a global basis. We further design a cue digging module (CDM), which utilizes local features to guide the learning trend of the model for highlighting the weak local responses. Finally, comprehensive experiments are carried out on two widely used datasets, i.e., CUB-200-2011 and ILSVRC, to verify the effectiveness of our method.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 872850e5-650e-4169-b20b-3021b41f7c0cCited by top-tier papers5
- Dynamic Prototype Mask for Occluded Person Re-IdentificationLei Tan, Pingyang Dai, Rongrong Ji, Yongjian WuACM MM 2022 · 94 citations
- Generative Prompt Model for Weakly Supervised Object LocalizationYuzhong Zhao, Qixiang Ye, Weijia Wu, Chunhua Shen et al.ICCV 2023 · 43 citations
- SEFormer: Structure Embedding Transformer for 3D Object DetectionXiaoyu Feng, Heming Du, Hehe Fan, Yueqi Duan et al.AAAI 2023 · 15 citations
- Coarse2Fine: Local Consistency Aware Re-prediction for Weakly Supervised Object LocalizationYixuan Pan, Yao Yao, Yichao Cao, Chongjin Chen et al.AAAI 2023 · 9 citations
- CAKE: Category Aware Knowledge Extraction for Open-Vocabulary Object DetectionShiyuan Ma, Donglin Qian, Kai Ye, Shengchuan ZhangAAAI 2025 · 8 citations
Builds on12
- Training data-efficient image transformers & distillation through attentionHugo Touvron, Matthieu Cord, Matthijs Douze, Francisco Massa et al.ICML 2021 · 8,974 citations
- CutMix: Regularization Strategy to Train Strong Classifiers With Localizable FeaturesSangdoo Yun, Dongyoon Han, Sanghyuk Chun, Seong Joon Oh et al.ICCV 2019 · 5,843 citations
- Tokens-to-Token ViT: Training Vision Transformers from Scratch on ImageNetLi Yuan, Yunpeng Chen, Tao Wang, Weihao Yu et al.ICCV 2021 · 2,462 citations
- TS-CAM: Token Semantic Coupled Attention Map for Weakly Supervised Object LocalizationWei Gao, Fang Wan, Xingjia Pan, Zhiliang Peng et al.ICCV 2021 · 260 citations
- DANet: Divergent Activation for Weakly Supervised Object LocalizationHaolan Xue, Chang Liu, Fang Wan, Jianbin Jiao et al.ICCV 2019 · 192 citations
Related papers
- Category-aware Allocation Transformer for Weakly Supervised Object LocalizationZhiwei Chen, Jinren Ding, Liujuan Cao, Yunhang Shen et al.ICCV 2023 · 15 citations
- Proxy Probing Decoder for Weakly Supervised Object Localization: A Baseline InvestigationJingyuan Xu, Hongtao Xie, Chuanbin Liu, Yongdong ZhangACM MM 2022 · 3 citations
- Spatial-Aware Token for Weakly Supervised Object LocalizationPingyu Wu, Wei Zhai, Yang Cao, Jiebo Luo et al.ICCV 2023 · 19 citations
- MoRe: Class Patch Attention Needs Regularization for Weakly Supervised Semantic SegmentationZhiwei Yang, Yucong Meng, Kexue Fu, Shuo Wang et al.AAAI 2025 · 14 citations
- LocLoc: Low-level Cues and Local-area Guides for Weakly Supervised Object LocalizationXinzi Cao, Xiawu Zheng, Yunhang Shen, Ke Li et al.ACM MM 2023 · 3 citations
