Negative Data Augmentation
Abhishek Sinha, Kumar Ayush, Jiaming Song, Burak Uzkent, Hongxia Jin, Stefano Ermon
摘要
Data augmentation is often used to enlarge datasets with synthetic samples generated in accordance with the underlying data distribution. To enable a wider range of augmentations, we explore negative data augmentation strategies (NDA) that intentionally create out-of-distribution samples. We show that such negative out-of-distribution samples provide information on the support of the data distribution, and can be leveraged for generative modeling and representation learning. We introduce a new GAN training objective where we use NDA as an additional source of synthetic data for the discriminator. We prove that under suitable conditions, optimizing the resulting objective still recovers the true data distribution but can directly bias the generator towards avoiding samples that lack the desired structure. Empirically, models trained with our method achieve improved conditional/unconditional image generation along with improved anomaly detection capabilities. Further, we incorporate the same negative data augmentation strategy in a contrastive learning framework for self-supervised representation learning on images and videos, achieving improved performance on downstream image classification, object detection, and action recognition tasks. These results suggest that prior knowledge on what does not constitute valid data is an effective form of weak supervision across a range of unsupervised learning tasks.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper18
- Robust Learning Meets Generative Models: Can Proxy Distributions Improve Adversarial Robustness?Vikash Sehwag, Saeed Mahloujifar, Tinashe Handina, Sihui Dai 等ICLR 2022 · 被引用 150 次
- Social NCE: Contrastive Learning of Socially-aware Motion RepresentationsYuejiang Liu, Qi Yan, Alexandre AlahiICCV 2021 · 被引用 118 次
- Understanding and Improving Robustness of Vision Transformers through Patch-based Negative AugmentationYao Qin, Chiyuan Zhang, Ting Chen, Balaji Lakshminarayanan 等NeurIPS 2022 · 被引用 68 次
- Improving Self-supervised Learning with Automated Unsupervised Outlier ArbitrationYu Wang, Jingyang Lin, Jingjing Zou, Yingwei Pan 等NeurIPS 2021 · 被引用 13 次
- Harnessing Out-Of-Distribution Examples via Augmenting Content and StyleZhuo Huang, Xiaobo Xia, Li Shen, Bo Han 等ICLR 2023 · 被引用 10 次
它引用的顶会 Paper5
- CutMix: Regularization Strategy to Train Strong Classifiers With Localizable FeaturesSangdoo Yun, Dongyoon Han, Sanghyuk Chun, Seong Joon Oh 等ICCV 2019 · 被引用 5,843 次
- Understanding the Limitations of Variational Mutual Information EstimatorsJiaming Song, Stefano ErmonICLR 2020 · 被引用 243 次
- A critical analysis of self-supervision, or what we can learn from a single imageYuki Markus Asano, Christian Rupprecht, Andrea VedaldiICLR 2020 · 被引用 152 次
- Difference-Seeking Generative Adversarial Network-Unseen Sample GenerationYi Lin Sung, Sung-Hsien Hsieh, Soo-Chang Pei, Chun-Shien LuICLR 2020 · 被引用 5 次
- Momentum Contrast for Unsupervised Visual Representation LearningKaiming He, Haoqi Fan, Yuxin Wu, Saining Xie 等CVPR 2020
相关 Paper
- Training GANs with Stronger Augmentations via Contrastive DiscriminatorJongheon Jeong, Jinwoo ShinICLR 2021 · 被引用 68 次
- Turning Waste into Wealth: Leveraging Low-Quality Samples for Enhancing Continuous Conditional Generative Adversarial NetworksXin Ding, Yongwei Wang, Zuheng XuAAAI 2024 · 被引用 4 次
- Self-Supervised Representation Learning via Neighborhood-Relational EncodingMohammad Sabokrou, Mohammad Khalooei, Ehsan AdeliICCV 2019 · 被引用 38 次
- Learning Transferable Negative Prompts for Out-of-Distribution DetectionTianqi Li, Guansong Pang, Xiao Bai, Wenjun Miao 等CVPR 2024
- VOS: Learning What You Don't Know by Virtual Outlier SynthesisXuefeng Du, Zhaoning Wang, Mu Cai, Yixuan LiICLR 2022 · 被引用 417 次
