Spectrum Random Masking for Generalization in Image-based Reinforcement Learning
Yangru Huang, Peixi Peng, Yifan Zhao, Guangyao Chen, Yonghong Tian
Abstract
Generalization in image-based reinforcement learning (RL) aims to learn a robust policy that could be applied directly on unseen visual environments, which is a challenging task since agents usually tend to overfit to their training environment. To handle this problem, a natural approach is to increase the data diversity by image based augmentations. However, different with most vision tasks such as classification and detection, RL tasks are not always invariant to spatial based augmentations due to the entanglement of environment dynamics and visual appearance. In this paper, we argue with two principles for augmentations in RL: First , the augmented observations should facilitate learning a universal policy, which is robust to various distribution shifts. Second , the augmented data should be invariant to the learning signals such as action and reward. Following these rules, we revisit image-based RL tasks from the view of frequency domain and propose a novel augmentation method, namely Spectrum Random Masking (SRM),which is able to help agents to learn the whole frequency spectrum of observation for coping with various distributions and compatible with the pre-collected action and reward corresponding to original observation. Extensive experiments conducted on DMControl Generalization Benchmark demonstrate the proposed SRM achieves the state-of-the-art performance with strong generalization potentials.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 397ca079-1d24-46c1-9e32-40593ae63a4bCited by top-tier papers11
- Opening the Sim-to-Real Door for Humanoid Pixel-to-Action Policy TransferHaoru Xue, Tairan He, Zi Wang, Qingwei Ben et al.CVPR 2026 · 35 citations
- Learning Better with Less: Effective Augmentation for Sample-Efficient Visual Reinforcement LearningGuozheng Ma, Linrui Zhang, Haoyu Wang, Lu Li et al.NeurIPS 2023 · 24 citations
- Focus On What Matters: Separated Models For Visual-Based RL GeneralizationDi Zhang, Bowen Lv, Hai Zhang, Feifan Yang et al.NeurIPS 2024 · 14 citations
- DMR: Decomposed Multi-Modality Representations for Frames and Events Fusion in Visual Reinforcement LearningHaoran Xu, Peixi Peng, Guang Tan, Yuan Li et al.CVPR 2024 · 5 citations
- Diffusion Guided Adaptive Augmentation for Generalization in Visual Reinforcement LearningJeong Woon Lee, Hyoseok HwangICCV 2025 · 3 citations
Builds on23
- CutMix: Regularization Strategy to Train Strong Classifiers With Localizable FeaturesSangdoo Yun, Dongyoon Han, Sanghyuk Chun, Seong Joon Oh et al.ICCV 2019 · 5,843 citations
- CURL: Contrastive Unsupervised Representations for Reinforcement LearningMichael Laskin, Aravind Srinivas, Pieter AbbeelICML 2020 · 1,261 citations
- Image Augmentation Is All You Need: Regularizing Deep Reinforcement Learning from PixelsDenis Yarats, Ilya Kostrikov, Rob FergusICLR 2021 · 911 citations
- Reinforcement Learning with Augmented DataMichael Laskin, Kimin Lee, Adam Stooke, Lerrel Pinto et al.NeurIPS 2020 · 833 citations
- Leveraging Procedural Generation to Benchmark Reinforcement LearningKarl Cobbe, Christopher Hesse, Jacob Hilton, John SchulmanICML 2020 · 685 citations
Related papers
- A Simple Framework for Generalization in Visual RL under Dynamic Scene PerturbationsWonil Song, Hyesong Choi, Kwanghoon Sohn, Dongbo MinNeurIPS 2024 · 6 citations
- PQDA: Policy-Aligned Q-Consistency Meets Decoupled Augmentation for Generalizable Visual RLYun Zhou, Yuqiang Wu, Chunyu TanAAAI 2026
- Unsupervised Visual Attention and Invariance for Reinforcement LearningXudong Wang, Long Lian, Stella X. YuCVPR 2021
- Pre-Trained Image Encoder for Generalizable Visual Reinforcement LearningZhecheng Yuan, Zhengrong Xue, Bo Yuan, Xueqian Wang et al.NeurIPS 2022 · 112 citations
- SECANT: Self-Expert Cloning for Zero-Shot Generalization of Visual PoliciesLinxi Fan, Guanzhi Wang, De-An Huang, Zhiding Yu et al.ICML 2021 · 73 citations
