A Data-Augmentation Is Worth A Thousand Samples: Analytical Moments And Sampling-Free Training
Randall Balestriero, Ishan Misra, Yann LeCun
摘要
Data-Augmentation (DA) is known to improve performance across tasks and datasets. We propose a method to theoretically analyze the effect of DA and study questions such as: how many augmented samples are needed to correctly estimate the information encoded by that DA? How does the augmentation policy impact the final parameters of a model? We derive several quantities in close-form, such as the expectation and variance of an image, loss, and model’s output under a given DA distribution. Up to our knowledge, we obtain the first explicit regularizer that corresponds to using DA during training for non-trivial transformations such as affine transformations, color jittering, or Gaussian blur. Those derivations open new avenues to quantify the benefits and limitations of DA. For example, given a loss at hand, we find that common DAs require tens of thousands of samples for the loss to be correctly estimated and for the model training to converge. We then show that for a training loss to have reduced variance under DA sampling, the model’s saliency map (gradient of the loss with respect to the model’s input) must align with the smallest eigenvector of the sample’s covariance matrix under the considered DA augmentation; this is exactly the quantity estimated and regularized by TangentProp. Those findings also hint at a possible explanation on why models tend to shift their focus from edges to textures when specific DAs are employed.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper8
- No Representation Rules Them All in Category DiscoverySagar Vaze, Andrea Vedaldi, Andrew ZissermanNeurIPS 2023 · 被引用 79 次
- Joint-Embedding vs Reconstruction: Provable Benefits of Latent Space Prediction for Self-Supervised LearningHugues Van Assel, Mark Ibrahim, Tommaso Biancalani, Aviv Regev 等NeurIPS 2025 · 被引用 39 次
- SF(DA)2: Source-free Domain Adaptation Through the Lens of Data AugmentationUiwon Hwang, Jonghyun Lee, Juhyeon Shin, Sungroh YoonICLR 2024 · 被引用 31 次
- How Much Data Are Augmentations Worth? An Investigation into Scaling Laws, Invariance, and Implicit RegularizationJonas Geiping, Micah Goldblum, Gowthami Somepalli, Ravid Shwartz-Ziv 等ICLR 2023 · 被引用 11 次
- Revisiting Data Augmentation in Deep Reinforcement LearningJianshu Hu, Yunpeng Jiang, Paul WengICLR 2024 · 被引用 9 次
它引用的顶会 Paper5
- CutMix: Regularization Strategy to Train Strong Classifiers With Localizable FeaturesSangdoo Yun, Dongyoon Han, Sanghyuk Chun, Seong Joon Oh 等ICCV 2019 · 被引用 5,843 次
- Barlow Twins: Self-Supervised Learning via Redundancy ReductionJure Zbontar, Li Jing, Ishan Misra, Yann LeCun 等ICML 2021 · 被引用 2,942 次
- Implicit Regularization in Deep Learning May Not Be Explainable by NormsNoam Razin, Nadav CohenNeurIPS 2020 · 被引用 178 次
- The Implicit and Explicit Regularization Effects of DropoutColin Wei, Sham M. Kakade, Tengyu MaICML 2020 · 被引用 129 次
- Self-Supervised Learning of Pretext-Invariant RepresentationsIshan Misra, Laurens van der MaatenCVPR 2020
相关 Paper
- Tradeoffs in Data Augmentation: An Empirical StudyRaphael Gontijo Lopes, Sylvia J. Smullin, Ekin Dogus Cubuk, Ethan DyerICLR 2021 · 被引用 72 次
- Data augmentation for deep learning based accelerated MRI reconstruction with limited dataZalan Fabian, Reinhard Heckel, Mahdi SoltanolkotabiICML 2021 · 被引用 60 次
- First-Order Manifold Data Augmentation for Regression LearningIlya Kaufman, Omri AzencotICML 2024 · 被引用 6 次
- KeepAugment: A Simple Information-Preserving Data Augmentation ApproachChengyue Gong, Dilin Wang, Meng Li, Vikas Chandra 等CVPR 2021
- A Group-Theoretic Framework for Data AugmentationShuxiao Chen, Edgar Dobriban, Jane H. LeeNeurIPS 2020 · 被引用 254 次
