MoVie: Visual Model-Based Policy Adaptation for View Generalization
Sizhe Yang, Yanjie Ze, Huazhe Xu
摘要
Visual Reinforcement Learning (RL) agents trained on limited views face significant challenges in generalizing their learned abilities to unseen views. This inherent difficulty is known as the problem of . In this work, we systematically categorize this fundamental problem into four distinct and highly challenging scenarios that closely resemble real-world situations. Subsequently, we propose a straightforward yet effective approach to enable successful adaptation of visual del-based policies for w generalization () during test time, without any need for explicit reward signals and any modification during training time. Our method demonstrates substantial advancements across all four scenarios encompassing a total of tasks sourced from DMControl, xArm, and Adroit, with a relative improvement of %, %, and % respectively. The superior results highlight the immense potential of our approach for real-world robotics applications. Videos are available at https://yangsizhe.github.io/MoVie/ .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- TD-MPC2: Scalable, Robust World Models for Continuous ControlNicklas Hansen, Hao Su, Xiaolong WangICLR 2024 · 被引用 388 次
- Learning an Actionable Discrete Diffusion Policy via Large-Scale Actionless Video Pre-TrainingHaoran He, Chenjia Bai, Ling Pan, Weinan Zhang 等NeurIPS 2024 · 被引用 38 次
- Smooth and Flexible Camera Movement Synthesis via Temporal Masked Generative ModelingChenghao Xu, Guangtao Lyu, Jiexi Yan, Muli Yang 等NeurIPS 2025 · 被引用 2 次
- Focus-Then-Reuse: Fast Adaptation in Visual Perturbation EnvironmentsJiahui Wang, Chao Chen, Jiacheng Xu, Zongzhang Zhang 等NeurIPS 2025 · 被引用 1 次
它引用的顶会 Paper16
- Test-Time Training with Self-Supervision for Generalization under Distribution ShiftsYu Sun, Xiaolong Wang, Zhuang Liu, John Miller 等ICML 2020 · 被引用 1,220 次
- Reinforcement Learning with Augmented DataMichael Laskin, Kimin Lee, Adam Stooke, Lerrel Pinto 等NeurIPS 2020 · 被引用 833 次
- Mastering Visual Continuous Control: Improved Data-Augmented Reinforcement LearningDenis Yarats, Rob Fergus, Alessandro Lazaric, Lerrel PintoICLR 2022 · 被引用 457 次
- Temporal Difference Learning for Model Predictive ControlNicklas Hansen, Hao Su, Xiaolong WangICML 2022 · 被引用 388 次
- Test-Time Training with Masked AutoencodersYossi Gandelsman, Yu Sun, Xinlei Chen, Alexei A. EfrosNeurIPS 2022 · 被引用 283 次
相关 Paper
- Self-Supervised Policy Adaptation during DeploymentNicklas Hansen, Rishabh Jangir, Yu Sun, Guillem Alenyà 等ICLR 2021 · 被引用 187 次
- Pre-Trained Image Encoder for Generalizable Visual Reinforcement LearningZhecheng Yuan, Zhengrong Xue, Bo Yuan, Xueqian Wang 等NeurIPS 2022 · 被引用 112 次
- Unsupervised Visual Attention and Invariance for Reinforcement LearningXudong Wang, Long Lian, Stella X. YuCVPR 2021
- Focus On What Matters: Separated Models For Visual-Based RL GeneralizationDi Zhang, Bowen Lv, Hai Zhang, Feifan Yang 等NeurIPS 2024 · 被引用 14 次
- Visual Adversarial Imitation Learning using Variational ModelsRafael Rafailov, Tianhe Yu, Aravind Rajeswaran, Chelsea FinnNeurIPS 2021 · 被引用 61 次
