SLACK: Stable Learning of Augmentations with Cold-Start and KL Regularization
Juliette Marrie, Michael Arbel, Diane Larlus, Julien Mairal
Abstract
Data augmentation is known to improve the generalization capabilities of neural networks, provided that the set of transformations is chosen with care, a selection often performed manually. Automatic data augmentation aims at automating this process. However, most recent approaches still rely on some prior information; they start from a small pool of manually-selected default transformations that are either used to pretrain the network or forced to be part of the policy learned by the automatic data augmentation algorithm. In this paper, we propose to directly learn the augmentation policy without leveraging such prior knowledge. The resulting bilevel optimization problem becomes more challenging due to the larger search space and the inherent instability of bilevel optimization algorithms. To mitigate these issues (i) we follow a successive cold-start strategy with a Kullback-Leibler regularization, and (ii) we parameterize magnitudes as continuous distributions. Our approach leads to competitive results on standard benchmarks despite a more challenging setting, and generalizes beyond natural images. 1
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 243204cc-b071-4e1d-b26f-27556a95e393Builds on9
- RandAugment: Practical Automated Data Augmentation with a Reduced Search SpaceEkin Dogus Cubuk, Barret Zoph, Jonathon Shlens, Quoc LeNeurIPS 2020 · 4,453 citations
- Moment Matching for Multi-Source Domain AdaptationXingchao Peng, Qinxun Bai, Xide Xia, Zijun Huang et al.ICCV 2019 · 2,239 citations
- In Search of Lost Domain GeneralizationIshaan Gulrajani, David Lopez-PazICLR 2021 · 1,416 citations
- TrivialAugment: Tuning-free Yet State-of-the-Art Data AugmentationSamuel G. Müller, Frank HutterICCV 2021 · 384 citations
- Addressing Model Vulnerability to Distributional Shifts Over Image Transformation SetsRiccardo Volpi, Vittorio MurinoICCV 2019 · 106 citations
Related papers
- Deep invariant networks with differentiable augmentation layersCédric Rommel, Thomas Moreau, Alexandre GramfortNeurIPS 2022 · 11 citations
- Deep AutoAugmentYu Zheng, Zhi Zhang, Shen Yan, Mi ZhangICLR 2022 · 32 citations
- Direct Differentiable Augmentation SearchAoming Liu, Zehao Huang, Zhiwu Huang, Naiyan WangICCV 2021 · 47 citations
- When to Learn What: Model-Adaptive Data Augmentation CurriculumChengkai Hou, Jieyu Zhang, Tianyi ZhouICCV 2023 · 25 citations
- Adversarial Auto-Augment with Label Preservation: A Representation Learning Principle Guided ApproachKaiwen Yang, Yanchao Sun, Jiahao Su, Fengxiang He et al.NeurIPS 2022 · 14 citations
