A Mechanistic Analysis of Sim-and-Real Co-Training in Generative Robot Policies
Yu Lei, Minghuan Liu, Abhiram Maddukuri, Zhenyu Jiang, Yuke Zhu
Abstract
Co-training, which combines limited in-domain real-world data with abundant surrogate data such as simulation or cross-embodiment robot data, is widely used for training generative robot policies. Despite its empirical success, the mechanisms that determine when and why co-training is effective remain poorly understood. We investigate the mechanism of sim-and-real co-training through theoretical analysis and empirical study, and identify two intrinsic effects governing performance. The first, "structured representation alignment", reflects a balance between cross-domain representation alignment and domain discernibility, and plays a primary role in downstream performance. The second, the "importance reweighting effect", arises from domain-dependent modulation of action weighting and operates at a secondary level. We validate these effects with controlled experiments on a toy model and extensive sim-and-sim and sim-and-real robot manipulation experiments. Our analysis offers a unified interpretation of recent co-training techniques and motivates a simple method that consistently improves upon prior approaches. More broadly, our aim is to examine the inner workings of co-training and to facilitate research in this direction.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext e836be67-cb9f-4010-bb66-af60e488b9fcBuilds on5
- Flow Straight and Fast: Learning to Generate and Transfer Data with Rectified FlowXingchao Liu, Chengyue Gong, Qiang LiuICLR 2023 · 75 citations
- Much Ado About Noising: Dispelling the Myths of Generative Robotic ControlChaoyi Pan, Giridharan Anantharaman, Nai-Chieh Huang, Claire Jin et al.ICLR 2026 · 51 citations
- EgoBridge: Domain Adaptation for Generalizable Imitation from Egocentric Human DataRyan Punamiya, Dhruv Patel, Patcharapong Aphiwetsa, Pranav Kuppili et al.NeurIPS 2025 · 40 citations
- Generalizable Domain Adaptation for Sim-and-Real Policy Co-TrainingShuo Cheng, Liqian Ma, Zhenyang Chen, Ajay Mandlekar et al.NeurIPS 2025 · 16 citations
- Latent Action Pretraining from VideosSeonghyeon Ye, Joel Jang, Byeongguk Jeon, Se June Joo et al.ICLR 2025
Related papers
- DataMIL: Selecting Data for Robot Imitation Learning with DatamodelsShivin Dass, Alaa Khaddaj, Logan Engstrom, Aleksander Madry et al.ICLR 2026 · 34 citations
- Guided Policy Optimization under Partial ObservabilityYueheng Li, Guangming Xie, Zongqing LuICLR 2026 · 4 citations
- Cross-modal Domain Adaptation for Cost-Efficient Visual Reinforcement LearningXiong-Hui Chen, Shengyi Jiang, Feng Xu, Zongzhang Zhang et al.NeurIPS 2021 · 15 citations
- Scalable and General Whole-Body Control for Cross-Humanoid LocomotionYufei Xue, Yunfeng Lin, Wentao Dong, Yang Tang et al.ICML 2026
- Learning Cross-Domain Correspondence for Control with Dynamics Cycle-ConsistencyQiang Zhang, Tete Xiao, Alexei A. Efros, Lerrel Pinto et al.ICLR 2021 · 73 citations
