Negative Data Augmentation
Abhishek Sinha, Kumar Ayush, Jiaming Song, Burak Uzkent, Hongxia Jin, Stefano Ermon
Abstract
Data augmentation is often used to enlarge datasets with synthetic samples generated in accordance with the underlying data distribution. To enable a wider range of augmentations, we explore negative data augmentation strategies (NDA) that intentionally create out-of-distribution samples. We show that such negative out-of-distribution samples provide information on the support of the data distribution, and can be leveraged for generative modeling and representation learning. We introduce a new GAN training objective where we use NDA as an additional source of synthetic data for the discriminator. We prove that under suitable conditions, optimizing the resulting objective still recovers the true data distribution but can directly bias the generator towards avoiding samples that lack the desired structure. Empirically, models trained with our method achieve improved conditional/unconditional image generation along with improved anomaly detection capabilities. Further, we incorporate the same negative data augmentation strategy in a contrastive learning framework for self-supervised representation learning on images and videos, achieving improved performance on downstream image classification, object detection, and action recognition tasks. These results suggest that prior knowledge on what does not constitute valid data is an effective form of weak supervision across a range of unsupervised learning tasks.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 0652c739-afcd-4c6c-a7da-5a16b953b24bCited by top-tier papers18
- Robust Learning Meets Generative Models: Can Proxy Distributions Improve Adversarial Robustness?Vikash Sehwag, Saeed Mahloujifar, Tinashe Handina, Sihui Dai et al.ICLR 2022 · 150 citations
- Social NCE: Contrastive Learning of Socially-aware Motion RepresentationsYuejiang Liu, Qi Yan, Alexandre AlahiICCV 2021 · 118 citations
- Understanding and Improving Robustness of Vision Transformers through Patch-based Negative AugmentationYao Qin, Chiyuan Zhang, Ting Chen, Balaji Lakshminarayanan et al.NeurIPS 2022 · 68 citations
- Improving Self-supervised Learning with Automated Unsupervised Outlier ArbitrationYu Wang, Jingyang Lin, Jingjing Zou, Yingwei Pan et al.NeurIPS 2021 · 13 citations
- Harnessing Out-Of-Distribution Examples via Augmenting Content and StyleZhuo Huang, Xiaobo Xia, Li Shen, Bo Han et al.ICLR 2023 · 10 citations
Builds on5
- CutMix: Regularization Strategy to Train Strong Classifiers With Localizable FeaturesSangdoo Yun, Dongyoon Han, Sanghyuk Chun, Seong Joon Oh et al.ICCV 2019 · 5,843 citations
- Understanding the Limitations of Variational Mutual Information EstimatorsJiaming Song, Stefano ErmonICLR 2020 · 243 citations
- A critical analysis of self-supervision, or what we can learn from a single imageYuki Markus Asano, Christian Rupprecht, Andrea VedaldiICLR 2020 · 152 citations
- Difference-Seeking Generative Adversarial Network-Unseen Sample GenerationYi Lin Sung, Sung-Hsien Hsieh, Soo-Chang Pei, Chun-Shien LuICLR 2020 · 5 citations
- Momentum Contrast for Unsupervised Visual Representation LearningKaiming He, Haoqi Fan, Yuxin Wu, Saining Xie et al.CVPR 2020
Related papers
- Training GANs with Stronger Augmentations via Contrastive DiscriminatorJongheon Jeong, Jinwoo ShinICLR 2021 · 68 citations
- Turning Waste into Wealth: Leveraging Low-Quality Samples for Enhancing Continuous Conditional Generative Adversarial NetworksXin Ding, Yongwei Wang, Zuheng XuAAAI 2024 · 4 citations
- Self-Supervised Representation Learning via Neighborhood-Relational EncodingMohammad Sabokrou, Mohammad Khalooei, Ehsan AdeliICCV 2019 · 38 citations
- Learning Transferable Negative Prompts for Out-of-Distribution DetectionTianqi Li, Guansong Pang, Xiao Bai, Wenjun Miao et al.CVPR 2024
- VOS: Learning What You Don't Know by Virtual Outlier SynthesisXuefeng Du, Zhaoning Wang, Mu Cai, Yixuan LiICLR 2022 · 417 citations
