Hierarchically Decoupled Imitation For Morphological Transfer
Donald J. Hejna III, Lerrel Pinto, Pieter Abbeel
Abstract
Learning long-range behaviors on complex high-dimensional agents is a fundamental problem in robot learning. For such tasks, we argue that transferring learned information from a morphologically simpler agent can massively improve the sample efficiency of a more complex one. To this end, we propose a hierarchical decoupling of policies into two parts: an independently learned low-level policy and a transferable high-level policy. To remedy poor transfer performance due to mismatch in morphologies, we contribute two key ideas. First, we show that incentivizing a complex agent's low-level to imitate a simpler agent's low-level significantly improves zero-shot high-level transfer. Second, we show that KL-regularized training of the high level stabilizes learning and prevents modecollapse. Finally, on a suite of publicly released navigation and manipulation environments, we demonstrate the applicability of hierarchical transfer on long-range tasks across morphologies. Our code and videos can be found at https: //sites.google.com/berkeley.edu/ morphology-transfer.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 565f3f15-0769-4e78-960a-35b868169d6eCited by top-tier papers15
- Cross-Domain Imitation Learning via Optimal TransportArnaud Fickinger, Samuel Cohen, Stuart Russell, Brandon AmosICLR 2022 · 65 citations
- Cross-Domain Policy Adaptation via Value-Guided Data FilteringKang Xu, Chenjia Bai, Xiaoteng Ma, Dong Wang et al.NeurIPS 2023 · 41 citations
- AnyMorph: Learning Transferable Polices By Inferring Agent MorphologyBrandon Trabucco, Mariano Phielipp, Glen BersethICML 2022 · 37 citations
- Cross-Domain Policy Adaptation by Capturing Representation MismatchJiafei Lyu, Chenjia Bai, Jingwen Yang, Zongqing Lu et al.ICML 2024 · 30 citations
- REvolveR: Continuous Evolutionary Models for Robot-to-robot Policy TransferXingyu Liu, Deepak Pathak, Kris KitaniICML 2022 · 24 citations
Builds on4
- Emergent Tool Use From Multi-Agent AutocurriculaBowen Baker, Ingmar Kanitscheider, Todor M. Markov, Yi Wu et al.ICLR 2020 · 751 citations
- Dynamics-Aware Unsupervised Discovery of SkillsArchit Sharma, Shixiang Gu, Sergey Levine, Vikash Kumar et al.ICLR 2020 · 475 citations
- Sub-policy Adaptation for Hierarchical Reinforcement LearningAlexander C. Li, Carlos Florensa, Ignasi Clavera, Pieter AbbeelICLR 2020 · 85 citations
- Discovering Motor Programs by Recomposing DemonstrationsTanmay Shankar, Shubham Tulsiani, Lerrel Pinto, Abhinav GuptaICLR 2020 · 58 citations
Related papers
- Toward Robust Long Range Policy TransferWei-Cheng Tseng, Jin-Siang Lin, Yao-Min Feng, Min SunAAAI 2021 · 8 citations
- Hierarchical Equivariant Policy via Frame TransferHaibo Zhao, Dian Wang, Yizhe Zhu, Xupeng Zhu et al.ICML 2025
- Distilling Morphology-Conditioned Hypernetworks for Efficient Universal Morphology ControlZheng Xiong, Risto Vuorio, Jacob Beck, Matthieu Zimmer et al.ICML 2024 · 8 citations
- Priors, Hierarchy, and Information Asymmetry for Skill Transfer in Reinforcement LearningSasha Salter, Kristian Hartikainen, Walter Goodwin, Ingmar PosnerICLR 2023 · 1 citation
- Composing Task-Agnostic Policies with Deep Reinforcement LearningAhmed Hussain Qureshi, Jacob J. Johnson, Yuzhe Qin, Taylor Henderson et al.ICLR 2020 · 35 citations
