KeepAugment: A Simple Information-Preserving Data Augmentation Approach
Chengyue Gong, Dilin Wang, Meng Li, Vikas Chandra, Qiang Liu
Abstract
Data augmentation (DA) is an essential technique for training state-of-the-art deep learning systems. In this paper, we empirically show that the standard data augmentation methods may introduce distribution shift and consequently hurt the performance on unaugmented data during inference. To alleviate this issue, we propose a simple yet effective approach, dubbed KeepAugment, to increase the fidelity of augmented images. The idea is to use the saliency map to detect important regions on the original images and preserve these informative regions during augmentation. This information-preserving strategy allows us to generate more faithful training examples. Empirically, we demonstrate that our method significantly improves upon a number of prior art data augmentation schemes, e.g. AutoAugment, Cutout, random erasing, achieving promising results on image classification, semi-supervised image classification, multi-view multi-camera tracking and object detection.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 55fc60e4-feab-4411-b417-1202a22959f2Cited by top-tier papers14
- Background-Mixed Augmentation for Weakly Supervised Change DetectionRui Huang, Ruofei Wang, Qing Guo, Jieda Wei et al.AAAI 2023 · 38 citations
- SageMix: Saliency-Guided Mixup for Point CloudsSanghyeok Lee, Minkyu Jeon, Injae Kim, Yunyang Xiong et al.NeurIPS 2022 · 37 citations
- IPMix: Label-Preserving Data Augmentation Method for Training Robust ClassifiersZhenglin Huang, Xiaoan Bao, Na Zhang, Qingqi Zhang et al.NeurIPS 2023 · 26 citations
- LEMON: Lossless model expansionYite Wang, Jiahao Su, Hanlin Lu, Cong Xie et al.ICLR 2024 · 25 citations
- Hyperbolic Feature Augmentation via Distribution Estimation and Infinite Sampling on ManifoldsZhi Gao, Yuwei Wu, Yunde Jia, Mehrtash HarandiNeurIPS 2022 · 21 citations
Builds on5
- CutMix: Regularization Strategy to Train Strong Classifiers With Localizable FeaturesSangdoo Yun, Dongyoon Han, Sanghyuk Chun, Seong Joon Oh et al.ICCV 2019 · 5,843 citations
- FixMatch: Simplifying Semi-Supervised Learning with Consistency and ConfidenceKihyuk Sohn, David Berthelot, Nicholas Carlini, Zizhao Zhang et al.NeurIPS 2020 · 5,129 citations
- Random Erasing Data AugmentationZhun Zhong, Liang Zheng, Guoliang Kang, Shaozi Li et al.AAAI 2020 · 4,134 citations
- AugMix: A Simple Data Processing Method to Improve Robustness and UncertaintyDan Hendrycks, Norman Mu, Ekin Dogus Cubuk, Barret Zoph et al.ICLR 2020 · 1,572 citations
- What It Thinks Is Important Is Important: Robustness Transfers Through Input GradientsAlvin Chan, Yi Tay, Yew-Soon OngCVPR 2020
Related papers
- SaliencyMix: A Saliency Guided Data Augmentation Strategy for Better RegularizationA. F. M. Shahab Uddin, Mst. Sirazam Monira, Wheemyung Shin, TaeChoong Chung et al.ICLR 2021 · 271 citations
- GuidedMixup: An Efficient Mixup Strategy Guided by Saliency MapsMinsoo Kang, Suhyun KimAAAI 2023 · 31 citations
- SelectAugment: Hierarchical Deterministic Sample Selection for Data AugmentationShiqi Lin, Zhizheng Zhang, Xin Li, Zhibo ChenAAAI 2023 · 13 citations
- SuperMix: Supervising the Mixing Data AugmentationAli Dabouei, Sobhan Soleymani, Fariborz Taherkhani, Nasser M. NasrabadiCVPR 2021
- ReMixMatch: Semi-Supervised Learning with Distribution Matching and Augmentation AnchoringDavid Berthelot, Nicholas Carlini, Ekin D. Cubuk, Alex Kurakin et al.ICLR 2020 · 469 citations
