Translating Robot Skills: Learning Unsupervised Skill Correspondences Across Robots
Tanmay Shankar, Yixin Lin, Aravind Rajeswaran, Vikash Kumar, Stuart Anderson, Jean Oh
摘要
In this paper, we explore how we can endow robots with the ability to learn correspondences between their own skills, and those of agents with different embodiments and in different domains than their own, in an entirely unsupervised manner. Our insight and premise is that agents with different embodiments use similar strategies (high-level skill sequences) to solve similar tasks. Based on this insight, we frame learning skill correspondences as a problem of matching distributions of sequences of skills across agents. We then present an unsupervised objective that encourages a learnt skill translation model to match these distributions across domains inspired by recent advances in unsupervised machine translation. Our approach is able to learn semantically meaningful correspondences between skills across multiple robot-robot and human-robot domain pairs, despite being completely unsupervised. Further, the learnt correspondences enable the transfer of task strategies across robots and domains. Dynamic visualization of our results can be found here: https://sites.google.com/view/ translatingrobotskills/home
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper8
- Dynamics-Aware Unsupervised Discovery of SkillsArchit Sharma, Shixiang Gu, Sergey Levine, Vikash Kumar 等ICLR 2020 · 被引用 475 次
- Skeleton-aware networks for deep motion retargetingKfir Aberman, Peizhuo Li, Dani Lischinski, Olga Sorkine-Hornung 等SIGGRAPH 2020 · 被引用 210 次
- State Alignment-based Imitation LearningFangchen Liu, Zhan Ling, Tongzhou Mu, Hao SuICLR 2020 · 被引用 103 次
- Domain Adaptive Imitation LearningKuno Kim, Yihong Gu, Jiaming Song, Shengjia Zhao 等ICML 2020 · 被引用 86 次
- Learning Robot Skills with Temporal Variational InferenceTanmay Shankar, Abhinav GuptaICML 2020 · 被引用 80 次
相关 Paper
- Unsupervised Domain Adaptation with Dynamics-Aware Rewards in Reinforcement LearningJinxin Liu, Hao Shen, Donglin Wang, Yachen Kang 等NeurIPS 2021 · 被引用 20 次
- Learning Cross-Domain Correspondence for Control with Dynamics Cycle-ConsistencyQiang Zhang, Tete Xiao, Alexei A. Efros, Lerrel Pinto 等ICLR 2021 · 被引用 73 次
- Unsupervised Skill Discovery for Learning Shared Structures across Changing EnvironmentsSang-Hyun Lee, Seung-Woo SeoICML 2023 · 被引用 6 次
- SkiLD: Unsupervised Skill Discovery Guided by Factor InteractionsZizhao Wang, Jiaheng Hu, Caleb Chuck, Stephen Chen 等NeurIPS 2024 · 被引用 15 次
- SemTra: A Semantic Skill Translator for Cross-Domain Zero-Shot Policy AdaptationSangwoo Shin, Minjong Yoo, Jeongwoo Lee, Honguk WooAAAI 2024 · 被引用 6 次
