AutoDO: Robust AutoAugment for Biased Data With Label Noise via Scalable Probabilistic Implicit Differentiation
Denis A. Gudovskiy, Luca Rigazio, Shun Ishizaka, Kazuki Kozuka, Sotaro Tsukizawa
摘要
AutoAugment [6] has sparked an interest in automated augmentation methods for deep learning models. These methods estimate image transformation policies for train data that improve generalization to test data. While recent papers evolved in the direction of decreasing policy search complexity, we show that those methods are not robust when applied to biased and noisy data. To overcome these limitations, we reformulate AutoAugment as a generalized automated dataset optimization (AutoDO) task that minimizes the distribution shift between test data and distorted train dataset. In our AutoDO model, we explicitly estimate a set of per-point hyperparameters to flexibly change distribution of train data. In particular, we include hyperparameters for augmentation, loss weights, and softlabels that are jointly estimated using implicit differentiation. We develop a theoretical probabilistic interpretation of this framework using Fisher information and show that its complexity scales linearly with the dataset size. Our experiments on SVHN, CIFAR-10/100, and ImageNet classification show up to 9.3% improvement for biased datasets with label noise compared to prior methods and, importantly, up to 36.6% gain for underrepresented SVHN classes 1 .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper6
- FedFixer: Mitigating Heterogeneous Label Noise in Federated LearningXinyuan Ji, Zhaowei Zhu, Wei Xi, Olga Gadyatskaya 等AAAI 2024 · 被引用 30 次
- Making Scalable Meta Learning PracticalSang Keun Choe, Sanket Vaibhav Mehta, Hwijeen Ahn, Willie Neiswanger 等NeurIPS 2023 · 被引用 28 次
- Hyperbolic Feature Augmentation via Distribution Estimation and Infinite Sampling on ManifoldsZhi Gao, Yuwei Wu, Yunde Jia, Mehrtash HarandiNeurIPS 2022 · 被引用 21 次
- CUDA: Curriculum of Data Augmentation for Long-tailed RecognitionSumyeong Ahn, Jongwoo Ko, Se-Young YunICLR 2023 · 被引用 15 次
- Efficient Scheduling of Data Augmentation for Deep Reinforcement LearningByungchan Ko, Jungseul OkNeurIPS 2022 · 被引用 6 次
它引用的顶会 Paper5
- CutMix: Regularization Strategy to Train Strong Classifiers With Localizable FeaturesSangdoo Yun, Dongyoon Han, Sanghyuk Chun, Seong Joon Oh 等ICCV 2019 · 被引用 5,843 次
- RandAugment: Practical Automated Data Augmentation with a Reduced Search SpaceEkin Dogus Cubuk, Barret Zoph, Jonathon Shlens, Quoc LeNeurIPS 2020 · 被引用 4,453 次
- Random Erasing Data AugmentationZhun Zhong, Liang Zheng, Guoliang Kang, Shaozi Li 等AAAI 2020 · 被引用 4,134 次
- Online Hyper-Parameter Learning for Auto-Augmentation StrategyChen Lin, Minghao Guo, Chuming Li, Xin Yuan 等ICCV 2019 · 被引用 92 次
- Deep Active Learning for Biased Datasets via Fisher Kernel Self-SupervisionDenis A. Gudovskiy, Alec Hodgkinson, Takuya Yamaguchi, Sotaro TsukizawaCVPR 2020
相关 Paper
- AdaAug: Learning Class- and Instance-adaptive Data Augmentation PoliciesTsz-Him Cheung, Dit-Yan YeungICLR 2022 · 被引用 31 次
- Adversarial AutoAugmentXinyu Zhang, Qiang Wang, Jian Zhang, Zhao ZhongICLR 2020 · 被引用 210 次
- MetaAugment: Sample-Aware Data Augmentation Policy LearningFengwei Zhou, Jiawei Li, Chuanlong Xie, Fei Chen 等AAAI 2021 · 被引用 35 次
- SLACK: Stable Learning of Augmentations with Cold-Start and KL RegularizationJuliette Marrie, Michael Arbel, Diane Larlus, Julien MairalCVPR 2023
- AdaTransform: Adaptive Data TransformationZhiqiang Tang, Xi Peng, Tingfeng Li, Yizhe Zhu 等ICCV 2019 · 被引用 20 次
