Learning Cross-Domain Correspondence for Control with Dynamics Cycle-Consistency
Qiang Zhang, Tete Xiao, Alexei A. Efros, Lerrel Pinto, Xiaolong Wang
Abstract
At the heart of many robotics problems is the challenge of learning correspondences across domains. For instance, imitation learning requires obtaining correspondence between humans and robots; sim-to-real requires correspondence between physics simulators and the real world; transfer learning requires correspondences between different robotics environments. This paper aims to learn correspondence across domains differing in representation (vision vs. internal state), physics parameters (mass and friction), and morphology (number of limbs). Importantly, correspondences are learned using unpaired and randomly collected data from the two domains. We propose dynamics cycles that align dynamic robot behavior across two domains using a cycle-consistency constraint. Once this correspondence is found, we can directly transfer the policy trained on one domain to the other, without needing any additional fine-tuning on the second domain. We perform experiments across a variety of problem domains, both in simulation and on real robot. Our framework is able to align uncalibrated monocular video of a real robot arm to dynamic state-action trajectories of a simulated arm without paired data. Video demonstrations of our results are available at: https://sjtuzq.github.io/cycle_dynamics.html .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 583fa24e-a30d-42e7-bb3f-4ac76e0c8a18Cited by top-tier papers14
- PlayVirtual: Augmenting Cycle-Consistent Virtual Trajectories for Reinforcement LearningTao Yu, Cuiling Lan, Wenjun Zeng, Mingxiao Feng et al.NeurIPS 2021 · 64 citations
- Cross-Domain Policy Adaptation via Value-Guided Data FilteringKang Xu, Chenjia Bai, Xiaoteng Ma, Dong Wang et al.NeurIPS 2023 · 41 citations
- Cross-Domain Policy Adaptation by Capturing Representation MismatchJiafei Lyu, Chenjia Bai, Jingwen Yang, Zongqing Lu et al.ICML 2024 · 30 citations
- CrossLoco: Human Motion Driven Control of Legged Robots via Guided Unsupervised Reinforcement LearningTianyu Li, Hyunyoung Jung, Matthew C. Gombolay, Yong Kwon Cho et al.ICLR 2024 · 14 citations
- CUP: Critic-Guided Policy ReuseJin Zhang, Siyuan Li, Chongjie ZhangNeurIPS 2022 · 11 citations
Builds on3
- What Makes for Good Views for Contrastive Learning?Yonglong Tian, Chen Sun, Ben Poole, Dilip Krishnan et al.NeurIPS 2020 · 1,631 citations
- DeceptionNet: Network-Driven Domain RandomizationSergey Zakharov, Wadim Kehl, Slobodan IlicICCV 2019 · 100 citations
- RL-CycleGAN: Reinforcement Learning Aware Simulation-to-RealKanishka Rao, Chris Harris, Alex Irpan, Sergey Levine et al.CVPR 2020
Related papers
- Cross-domain Imitation from ObservationsDripta S. Raychaudhuri, Sujoy Paul, Jeroen van Baar, Amit K. Roy-ChowdhuryICML 2021 · 54 citations
- Translating Robot Skills: Learning Unsupervised Skill Correspondences Across RobotsTanmay Shankar, Yixin Lin, Aravind Rajeswaran, Vikash Kumar et al.ICML 2022 · 10 citations
- Domain Adaptive Imitation Learning with Visual ObservationSungho Choi, Seungyul Han, Woojun Kim, Jongseong Chae et al.NeurIPS 2023 · 15 citations
- EgoBridge: Domain Adaptation for Generalizable Imitation from Egocentric Human DataRyan Punamiya, Dhruv Patel, Patcharapong Aphiwetsa, Pranav Kuppili et al.NeurIPS 2025 · 40 citations
- REvolveR: Continuous Evolutionary Models for Robot-to-robot Policy TransferXingyu Liu, Deepak Pathak, Kris KitaniICML 2022 · 24 citations
