Learning Cross-Domain Correspondence for Control with Dynamics Cycle-Consistency
Qiang Zhang, Tete Xiao, Alexei A. Efros, Lerrel Pinto, Xiaolong Wang
摘要
At the heart of many robotics problems is the challenge of learning correspondences across domains. For instance, imitation learning requires obtaining correspondence between humans and robots; sim-to-real requires correspondence between physics simulators and the real world; transfer learning requires correspondences between different robotics environments. This paper aims to learn correspondence across domains differing in representation (vision vs. internal state), physics parameters (mass and friction), and morphology (number of limbs). Importantly, correspondences are learned using unpaired and randomly collected data from the two domains. We propose dynamics cycles that align dynamic robot behavior across two domains using a cycle-consistency constraint. Once this correspondence is found, we can directly transfer the policy trained on one domain to the other, without needing any additional fine-tuning on the second domain. We perform experiments across a variety of problem domains, both in simulation and on real robot. Our framework is able to align uncalibrated monocular video of a real robot arm to dynamic state-action trajectories of a simulated arm without paired data. Video demonstrations of our results are available at: https://sjtuzq.github.io/cycle_dynamics.html .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper14
- PlayVirtual: Augmenting Cycle-Consistent Virtual Trajectories for Reinforcement LearningTao Yu, Cuiling Lan, Wenjun Zeng, Mingxiao Feng 等NeurIPS 2021 · 被引用 64 次
- Cross-Domain Policy Adaptation via Value-Guided Data FilteringKang Xu, Chenjia Bai, Xiaoteng Ma, Dong Wang 等NeurIPS 2023 · 被引用 41 次
- Cross-Domain Policy Adaptation by Capturing Representation MismatchJiafei Lyu, Chenjia Bai, Jingwen Yang, Zongqing Lu 等ICML 2024 · 被引用 30 次
- CrossLoco: Human Motion Driven Control of Legged Robots via Guided Unsupervised Reinforcement LearningTianyu Li, Hyunyoung Jung, Matthew C. Gombolay, Yong Kwon Cho 等ICLR 2024 · 被引用 14 次
- CUP: Critic-Guided Policy ReuseJin Zhang, Siyuan Li, Chongjie ZhangNeurIPS 2022 · 被引用 11 次
它引用的顶会 Paper3
- What Makes for Good Views for Contrastive Learning?Yonglong Tian, Chen Sun, Ben Poole, Dilip Krishnan 等NeurIPS 2020 · 被引用 1,631 次
- DeceptionNet: Network-Driven Domain RandomizationSergey Zakharov, Wadim Kehl, Slobodan IlicICCV 2019 · 被引用 100 次
- RL-CycleGAN: Reinforcement Learning Aware Simulation-to-RealKanishka Rao, Chris Harris, Alex Irpan, Sergey Levine 等CVPR 2020
相关 Paper
- Cross-domain Imitation from ObservationsDripta S. Raychaudhuri, Sujoy Paul, Jeroen van Baar, Amit K. Roy-ChowdhuryICML 2021 · 被引用 54 次
- Translating Robot Skills: Learning Unsupervised Skill Correspondences Across RobotsTanmay Shankar, Yixin Lin, Aravind Rajeswaran, Vikash Kumar 等ICML 2022 · 被引用 10 次
- Domain Adaptive Imitation Learning with Visual ObservationSungho Choi, Seungyul Han, Woojun Kim, Jongseong Chae 等NeurIPS 2023 · 被引用 15 次
- EgoBridge: Domain Adaptation for Generalizable Imitation from Egocentric Human DataRyan Punamiya, Dhruv Patel, Patcharapong Aphiwetsa, Pranav Kuppili 等NeurIPS 2025 · 被引用 40 次
- REvolveR: Continuous Evolutionary Models for Robot-to-robot Policy TransferXingyu Liu, Deepak Pathak, Kris KitaniICML 2022 · 被引用 24 次
