Revisiting Data Augmentation in Deep Reinforcement Learning
Jianshu Hu, Yunpeng Jiang, Paul Weng
Abstract
Various data augmentation techniques have been recently proposed in imagebased deep reinforcement learning (DRL). Although they empirically demonstrate the effectiveness of data augmentation for improving sample efficiency or generalization, which technique should be preferred is not always clear. To tackle this question, we analyze existing methods to better understand them and to uncover how they are connected. Notably, by expressing the variance of the Q-targets and that of the empirical actor/critic losses of these methods, we can analyze the effects of their different components and compare them. We furthermore formulate an explanation about how these methods may be affected by choosing different data augmentation transformations in calculating the target Q-values. This analysis suggests recommendations on how to exploit data augmentation in a more principled way. In addition, we include a regularization term called tangent prop, previously proposed in computer vision, but whose adaptation to DRL is novel to the best of our knowledge. We evaluate our proposition 1 and validate our analysis in several domains. Compared to different relevant baselines, we demonstrate that it achieves state-of-the-art performance in most environments and shows higher sample efficiency and better generalization ability in some complex environments.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 7809b702-d558-4fbd-bb2a-7fe7eb3fda87Cited by top-tier papers3
- Time Reversal Symmetry for Efficient Robotic Manipulations in Deep Reinforcement LearningYunpeng Jiang, Jianshu Hu, Paul Weng, Yutong BanNeurIPS 2025 · 1 citation
- Episodic Memory-Guided Controllable Experience Synthesis for Reinforcement LearningXiao Ma, Tian Li, Wu-Jun LiICML 2026
- TSTM: Temporal Segmentation for Task-relevant Mask in Visual Reinforcement Learning GeneralizationWeicheng Du, Wenjia Meng, Zhengzhe Zhang, Yilong Yin et al.CVPR 2026
Builds on11
- Bootstrap Your Own Latent - A New Approach to Self-Supervised LearningJean-Bastien Grill, Florian Strub, Florent Altché, Corentin Tallec et al.NeurIPS 2020 · 9,171 citations
- CURL: Contrastive Unsupervised Representations for Reinforcement LearningMichael Laskin, Aravind Srinivas, Pieter AbbeelICML 2020 · 1,261 citations
- Image Augmentation Is All You Need: Regularizing Deep Reinforcement Learning from PixelsDenis Yarats, Ilya Kostrikov, Rob FergusICLR 2021 · 911 citations
- Reinforcement Learning with Augmented DataMichael Laskin, Kimin Lee, Adam Stooke, Lerrel Pinto et al.NeurIPS 2020 · 833 citations
- Mastering Visual Continuous Control: Improved Data-Augmented Reinforcement LearningDenis Yarats, Rob Fergus, Alessandro Lazaric, Lerrel PintoICLR 2022 · 457 citations
Related papers
- A Data-Augmentation Is Worth A Thousand Samples: Analytical Moments And Sampling-Free TrainingRandall Balestriero, Ishan Misra, Yann LeCunNeurIPS 2022 · 23 citations
- Automatic Data Augmentation for Generalization in Reinforcement LearningRoberta Raileanu, Maxwell Goldstein, Denis Yarats, Ilya Kostrikov et al.NeurIPS 2021 · 143 citations
- Functional Regularization for Reinforcement Learning via Learned Fourier FeaturesAlexander C. Li, Deepak PathakNeurIPS 2021 · 27 citations
- Learning Better with Less: Effective Augmentation for Sample-Efficient Visual Reinforcement LearningGuozheng Ma, Linrui Zhang, Haoyu Wang, Lu Li et al.NeurIPS 2023 · 24 citations
- Reinforcement Learning with Euclidean Data Augmentation for State-Based Continuous ControlJinzhu Luo, Dingyang Chen, Qi ZhangNeurIPS 2024 · 5 citations
