A Mechanistic Analysis of Sim-and-Real Co-Training in Generative Robot Policies
Yu Lei, Minghuan Liu, Abhiram Maddukuri, Zhenyu Jiang, Yuke Zhu
摘要
Co-training, which combines limited in-domain real-world data with abundant surrogate data such as simulation or cross-embodiment robot data, is widely used for training generative robot policies. Despite its empirical success, the mechanisms that determine when and why co-training is effective remain poorly understood. We investigate the mechanism of sim-and-real co-training through theoretical analysis and empirical study, and identify two intrinsic effects governing performance. The first, "structured representation alignment", reflects a balance between cross-domain representation alignment and domain discernibility, and plays a primary role in downstream performance. The second, the "importance reweighting effect", arises from domain-dependent modulation of action weighting and operates at a secondary level. We validate these effects with controlled experiments on a toy model and extensive sim-and-sim and sim-and-real robot manipulation experiments. Our analysis offers a unified interpretation of recent co-training techniques and motivates a simple method that consistently improves upon prior approaches. More broadly, our aim is to examine the inner workings of co-training and to facilitate research in this direction.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper5
- Flow Straight and Fast: Learning to Generate and Transfer Data with Rectified FlowXingchao Liu, Chengyue Gong, Qiang LiuICLR 2023 · 被引用 75 次
- Much Ado About Noising: Dispelling the Myths of Generative Robotic ControlChaoyi Pan, Giridharan Anantharaman, Nai-Chieh Huang, Claire Jin 等ICLR 2026 · 被引用 51 次
- EgoBridge: Domain Adaptation for Generalizable Imitation from Egocentric Human DataRyan Punamiya, Dhruv Patel, Patcharapong Aphiwetsa, Pranav Kuppili 等NeurIPS 2025 · 被引用 40 次
- Generalizable Domain Adaptation for Sim-and-Real Policy Co-TrainingShuo Cheng, Liqian Ma, Zhenyang Chen, Ajay Mandlekar 等NeurIPS 2025 · 被引用 16 次
- Latent Action Pretraining from VideosSeonghyeon Ye, Joel Jang, Byeongguk Jeon, Se June Joo 等ICLR 2025
相关 Paper
- DataMIL: Selecting Data for Robot Imitation Learning with DatamodelsShivin Dass, Alaa Khaddaj, Logan Engstrom, Aleksander Madry 等ICLR 2026 · 被引用 34 次
- Guided Policy Optimization under Partial ObservabilityYueheng Li, Guangming Xie, Zongqing LuICLR 2026 · 被引用 4 次
- Cross-modal Domain Adaptation for Cost-Efficient Visual Reinforcement LearningXiong-Hui Chen, Shengyi Jiang, Feng Xu, Zongzhang Zhang 等NeurIPS 2021 · 被引用 15 次
- Scalable and General Whole-Body Control for Cross-Humanoid LocomotionYufei Xue, Yunfeng Lin, Wentao Dong, Yang Tang 等ICML 2026
- Learning Cross-Domain Correspondence for Control with Dynamics Cycle-ConsistencyQiang Zhang, Tete Xiao, Alexei A. Efros, Lerrel Pinto 等ICLR 2021 · 被引用 73 次
