TdAttenMix: Top-Down Attention Guided Mixup
Zhiming Wang, Lin Gu, Feng Lu
Abstract
CutMix is a data augmentation strategy that cuts and pastes image patches to mixup training data. Existing methods pick either random or salient areas which are often inconsistent to labels, thus misguiding the training model. By our knowledge, we integrate human gaze to guide cutmix for the first time. Since human attention is driven by both high-level recognition and low-level clues, we propose a controllable Top-down Attention Guided Module to obtain a general artificial attention which balances top-down and bottom-up attention. The proposed TdATttenMix then picks the patches and adjust the label mixing ratio that focuses on regions relevant to the current label. Experimental results demonstrate that our TdAttenMix outperforms existing state-of-the-art mixup methods across eight different benchmarks. Additionally, we introduce a new metric based on the human gaze and use this metric to investigate the issue of image-label inconsistency.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 6e2f66ea-015d-464a-afc1-dcc2b5619579Cited by top-tier papers1
Ask how each one uses itBuilds on16
- CutMix: Regularization Strategy to Train Strong Classifiers With Localizable FeaturesSangdoo Yun, Dongyoon Han, Sanghyuk Chun, Seong Joon Oh et al.ICCV 2019 · 5,843 citations
- CrossViT: Cross-Attention Multi-Scale Vision Transformer for Image ClassificationChun-Fu (Richard) Chen, Quanfu Fan, Rameswar PandaICCV 2021 · 2,072 citations
- Intriguing Properties of Vision TransformersMuzammal Naseer, Kanchana Ranasinghe, Salman Khan, Munawar Hayat et al.NeurIPS 2021 · 863 citations
- Puzzle Mix: Exploiting Saliency and Local Statistics for Optimal MixupJang-Hyun Kim, Wonho Choo, Hyun Oh SongICML 2020 · 457 citations
- SaliencyMix: A Saliency Guided Data Augmentation Strategy for Better RegularizationA. F. M. Shahab Uddin, Mst. Sirazam Monira, Wheemyung Shin, TaeChoong Chung et al.ICLR 2021 · 271 citations
Related papers
- SMMix: Self-Motivated Image Mixing for Vision TransformersMengzhao Chen, Mingbao Lin, Zhihang Lin, Yuxin Zhang et al.ICCV 2023 · 15 citations
- MixPro: Data Augmentation with MaskMix and Progressive Attention Labeling for Vision TransformerQihao Zhao, Yangyu Huang, Wei Hu, Fan Zhang et al.ICLR 2023 · 3 citations
- StyleMix: Separating Content and Style for Enhanced Data AugmentationMinui Hong, Jinwoo Choi, Gunhee KimCVPR 2021
- GuidedMixup: An Efficient Mixup Strategy Guided by Saliency MapsMinsoo Kang, Suhyun KimAAAI 2023 · 31 citations
- TokenMixup: Efficient Attention-guided Token-level Data Augmentation for TransformersHyeong Kyu Choi, Joonmyung Choi, Hyunwoo J. KimNeurIPS 2022 · 49 citations
