Anchor Data Augmentation
Nora Schneider, Shirin Goshtasbpour, Fernando Pérez-Cruz
Abstract
We propose a novel algorithm for data augmentation in nonlinear over-parametrized regression. Our data augmentation algorithm borrows from the literature on causality and extends the recently proposed Anchor regression (AR) method for data augmentation, which is in contrast to the current state-of-the-art domain-agnostic solutions that rely on the Mixup literature. Our Anchor Data Augmentation (ADA) uses several replicas of the modified samples in AR to provide more training examples, leading to more robust regression predictions. We apply ADA to linear and nonlinear regression problems using neural networks. ADA is competitive with state-of-the-art C-Mixup solutions. 1
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers7
- First-Order Manifold Data Augmentation for Regression LearningIlya Kaufman, Omri AzencotICML 2024 · 6 citations
- RC-Mixup: A Data Augmentation Strategy against Noisy Data for Regression TasksSeonghyeon Hwang, Minsu Kim, Steven Euijong WhangKDD 2024 · 4 citations
- ReAugment: Targeted Few-Shot Time Series Augmentation via Model Zoo-Guided Reinforcement LearningHaochen Yuan, Yutong Wang, Yihong Chen, Yunbo Wang et al.ICML 2026 · 3 citations
- Denoising Mixup for RegressionZhengzhang Hou, Zhanshan Li, Yanbo Liu, Geoff Nitschke et al.AAAI 2026
- Semi-Supervised Regression by Preserving Ranking Relationships Between Close Unlabeled SamplesXiming Li, Jiaxuan Jiang, Changchun Li, You Lu et al.AAAI 2026
Builds on10
- CutMix: Regularization Strategy to Train Strong Classifiers With Localizable FeaturesSangdoo Yun, Dongyoon Han, Sanghyuk Chun, Seong Joon Oh et al.ICCV 2019 · 5,843 citations
- RandAugment: Practical Automated Data Augmentation with a Reduced Search SpaceEkin Dogus Cubuk, Barret Zoph, Jonathon Shlens, Quoc LeNeurIPS 2020 · 4,453 citations
- WILDS: A Benchmark of in-the-Wild Distribution ShiftsPang Wei Koh, Shiori Sagawa, Henrik Marklund, Sang Michael Xie et al.ICML 2021 · 1,773 citations
- Puzzle Mix: Exploiting Saliency and Local Statistics for Optimal MixupJang-Hyun Kim, Wonho Choo, Hyun Oh SongICML 2020 · 457 citations
- TrivialAugment: Tuning-free Yet State-of-the-Art Data AugmentationSamuel G. Müller, Frank HutterICCV 2021 · 384 citations
Related papers
- Counterfactual Residual Data Augmentation for RegressionHossein Mohebbi, Oliver Schulte, Ke Li, Pascal PoupartICML 2026
- C-Mixup: Improving Generalization in RegressionHuaxiu Yao, Yiping Wang, Linjun Zhang, James Y. Zou et al.NeurIPS 2022 · 106 citations
- Improving Generalization in Reinforcement Learning with Mixture RegularizationKaixin Wang, Bingyi Kang, Jie Shao, Jiashi FengNeurIPS 2020 · 143 citations
- Nonlinear Mixup: Out-Of-Manifold Data Augmentation for Text ClassificationHongyu GuoAAAI 2020 · 124 citations
- Semi-Supervised Graph Imbalanced RegressionGang Liu, Tong Zhao, Eric Inae, Tengfei Luo et al.KDD 2023 · 20 citations
