MoVie: Visual Model-Based Policy Adaptation for View Generalization
Sizhe Yang, Yanjie Ze, Huazhe Xu
Abstract
Visual Reinforcement Learning (RL) agents trained on limited views face significant challenges in generalizing their learned abilities to unseen views. This inherent difficulty is known as the problem of . In this work, we systematically categorize this fundamental problem into four distinct and highly challenging scenarios that closely resemble real-world situations. Subsequently, we propose a straightforward yet effective approach to enable successful adaptation of visual del-based policies for w generalization () during test time, without any need for explicit reward signals and any modification during training time. Our method demonstrates substantial advancements across all four scenarios encompassing a total of tasks sourced from DMControl, xArm, and Adroit, with a relative improvement of %, %, and % respectively. The superior results highlight the immense potential of our approach for real-world robotics applications. Videos are available at https://yangsizhe.github.io/MoVie/ .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext be0f1ccf-18c8-4fcf-a42f-d7909664f303Cited by top-tier papers4
- TD-MPC2: Scalable, Robust World Models for Continuous ControlNicklas Hansen, Hao Su, Xiaolong WangICLR 2024 · 388 citations
- Learning an Actionable Discrete Diffusion Policy via Large-Scale Actionless Video Pre-TrainingHaoran He, Chenjia Bai, Ling Pan, Weinan Zhang et al.NeurIPS 2024 · 38 citations
- Smooth and Flexible Camera Movement Synthesis via Temporal Masked Generative ModelingChenghao Xu, Guangtao Lyu, Jiexi Yan, Muli Yang et al.NeurIPS 2025 · 2 citations
- Focus-Then-Reuse: Fast Adaptation in Visual Perturbation EnvironmentsJiahui Wang, Chao Chen, Jiacheng Xu, Zongzhang Zhang et al.NeurIPS 2025 · 1 citation
Builds on16
- Test-Time Training with Self-Supervision for Generalization under Distribution ShiftsYu Sun, Xiaolong Wang, Zhuang Liu, John Miller et al.ICML 2020 · 1,220 citations
- Reinforcement Learning with Augmented DataMichael Laskin, Kimin Lee, Adam Stooke, Lerrel Pinto et al.NeurIPS 2020 · 833 citations
- Mastering Visual Continuous Control: Improved Data-Augmented Reinforcement LearningDenis Yarats, Rob Fergus, Alessandro Lazaric, Lerrel PintoICLR 2022 · 457 citations
- Temporal Difference Learning for Model Predictive ControlNicklas Hansen, Hao Su, Xiaolong WangICML 2022 · 388 citations
- Test-Time Training with Masked AutoencodersYossi Gandelsman, Yu Sun, Xinlei Chen, Alexei A. EfrosNeurIPS 2022 · 283 citations
Related papers
- Self-Supervised Policy Adaptation during DeploymentNicklas Hansen, Rishabh Jangir, Yu Sun, Guillem Alenyà et al.ICLR 2021 · 187 citations
- Pre-Trained Image Encoder for Generalizable Visual Reinforcement LearningZhecheng Yuan, Zhengrong Xue, Bo Yuan, Xueqian Wang et al.NeurIPS 2022 · 112 citations
- Unsupervised Visual Attention and Invariance for Reinforcement LearningXudong Wang, Long Lian, Stella X. YuCVPR 2021
- Focus On What Matters: Separated Models For Visual-Based RL GeneralizationDi Zhang, Bowen Lv, Hai Zhang, Feifan Yang et al.NeurIPS 2024 · 14 citations
- Visual Adversarial Imitation Learning using Variational ModelsRafael Rafailov, Tianhe Yu, Aravind Rajeswaran, Chelsea FinnNeurIPS 2021 · 61 citations
