Know Thyself: Transferable Visual Control Policies Through Robot-Awareness
Edward S. Hu, Kun Huang, Oleh Rybkin, Dinesh Jayaraman
摘要
Training visual control policies from scratch on a new robot typically requires generating large amounts of robot-specific data. How might we leverage data previously collected on another robot to reduce or even completely remove this need for robot-specific data? We propose a "robot-aware control" paradigm that achieves this by exploiting readily available knowledge about the robot. We then instantiate this in a robot-aware model-based RL policy by training modular dynamics models that couple a transferable, robot-aware world dynamics module with a robot-specific, potentially analytical, robot dynamics module. This also enables us to set up visual planning costs that separately consider the robot agent and the world. Our experiments on tabletop manipulation tasks with simulated and real robots demonstrate that these plug-in improvements dramatically boost the transferability of visual model-based RL policies, even permitting zero-shot transfer of visual manipulation skills onto new robots. Project website: https: //www.seas.upenn.edu/ hued/rac
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper7
- Privileged Sensing Scaffolds Reinforcement LearningEdward S. Hu, James Springer, Oleh Rybkin, Dinesh JayaramanICLR 2024 · 被引用 21 次
- PEAC: Unsupervised Pre-training for Cross-Embodiment Reinforcement LearningChengyang Ying, Zhongkai Hao, Xinning Zhou, Xuezhou Xu 等NeurIPS 2024 · 被引用 14 次
- Efficient RL via Disentangled Environment and Agent RepresentationsKevin Gmelin, Shikhar Bahl, Russell Mendonca, Deepak PathakICML 2023 · 被引用 14 次
- OXE-AugE: A Large-Scale Robot Augmentation of OXE for Scaling Cross-Embodiment Policy LearningGuanhua Ji, Harsha Polavaram, Lawrence Yunliang Chen, Sandeep Bajamahal 等ICML 2026 · 被引用 14 次
- GraspGen-X: Cross-Embodiment 6-DOF Diffusion-based GraspingBeining Han, Yu-Wei Chao, Erwin Coumans, Clemens Eppner 等CVPR 2026 · 被引用 7 次
它引用的顶会 Paper8
- Hierarchical Foresight: Self-Supervised Learning of Long-Horizon Tasks via Visual Subgoal GenerationSuraj Nair, Chelsea FinnICLR 2020 · 被引用 152 次
- Goal-Aware Prediction: Learning to Model What MattersSuraj Nair, Silvio Savarese, Chelsea FinnICML 2020 · 被引用 71 次
- Model-Based Visual Planning with Self-Supervised Functional DistancesStephen Tian, Suraj Nair, Frederik Ebert, Sudeep Dasari 等ICLR 2021 · 被引用 69 次
- Cross-domain Imitation from ObservationsDripta S. Raychaudhuri, Sujoy Paul, Jeroen van Baar, Amit K. Roy-ChowdhuryICML 2021 · 被引用 54 次
- Model-Based Reinforcement Learning via Latent-Space CollocationOleh Rybkin, Chuning Zhu, Anusha Nagabandi, Kostas Daniilidis 等ICML 2021 · 被引用 46 次
相关 Paper
- AnyBimanual: Transferring Unimanual Policy for General Bimanual ManipulationGuanxing Lu, Tengbo Yu, Haoyuan Deng, Season Si Chen 等ICCV 2025 · 被引用 2 次
- On the Feasibility of Cross-Task Transfer with Model-Based Reinforcement LearningYifan Xu, Nicklas Hansen, Zirui Wang, Yung-Chieh Chan 等ICLR 2023 · 被引用 3 次
- Learning Temporally AbstractWorld Models without Online ExperimentationBenjamin Freed, Siddarth Venkatraman, Guillaume Adrien Sartoretti, Jeff Schneider 等ICML 2023 · 被引用 7 次
- Augmented World Models Facilitate Zero-Shot Dynamics Generalization From a Single Offline EnvironmentPhilip J. Ball, Cong Lu, Jack Parker-Holder, Stephen J. RobertsICML 2021 · 被引用 55 次
- STRAP: Robot Sub-Trajectory Retrieval for Augmented Policy LearningMarius Memmel, Jacob Berg, Bingqing Chen, Abhishek Gupta 等ICLR 2025
