CoMic: Complementary Task Learning & Mimicry for Reusable Skills
Leonard Hasenclever, Fabio Pardo, Raia Hadsell, Nicolas Heess, Josh Merel
摘要
Learning to control complex bodies and reuse learned behaviors is a longstanding challenge in continuous control. We study the problem of learning reusable humanoid skills by imitating motion capture data and joint training with complementary tasks. We show that it is possible to learn reusable skills through reinforcement learning on 50 times more motion capture data than prior work. We systematically compare a variety of different network architectures across different data regimes both in terms of imitation performance as well as transfer to challenging locomotion tasks. Finally we show that it is possible to interleave the motion capture tracking with training on complementary tasks, enriching the resulting skill space, and enabling the reuse of skills not well covered by the motion capture data such as getting up from the ground or catching a ball.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper17
- Perpetual Humanoid Control for Real-time Simulated AvatarsZhengyi Luo, Jinkun Cao, Alexander Winkler, Kris Kitani 等ICCV 2023 · 被引用 256 次
- ASE: large-scale reusable adversarial skill embeddings for physically simulated charactersXue Bin Peng, Yunrong Guo, Lina Halper, Sergey Levine 等SIGGRAPH 2022 · 被引用 217 次
- Universal Humanoid Motion Representations for Physics-Based ControlZhengyi Luo, Jinkun Cao, Josh Merel, Alexander Winkler 等ICLR 2024 · 被引用 125 次
- Omnigrasp: Grasping Diverse Objects with Simulated HumanoidsZhengyi Luo, Jinkun Cao, Sammy Christen, Alexander Winkler 等NeurIPS 2024 · 被引用 66 次
- PMP: Learning to Physically Interact with Environments using Part-wise Motion PriorsJinseok Bae, Jungdam Won, Donggeun Lim, Cheol-Hui Min 等SIGGRAPH 2023 · 被引用 27 次
它引用的顶会 Paper3
- Catch & Carry: reusable neural controllers for vision-guided whole-body tasksJosh Merel, Saran Tunyasuvunakool, Arun Ahuja, Yuval Tassa 等SIGGRAPH 2020 · 被引用 103 次
- A distributional view on multi-objective policy optimizationAbbas Abdolmaleki, Sandy H. Huang, Leonard Hasenclever, Michael Neunert 等ICML 2020 · 被引用 93 次
- The Variational Bandwidth Bottleneck: Stochastic Evaluation on an Information BudgetAnirudh Goyal, Yoshua Bengio, Matthew M. Botvinick, Sergey LevineICLR 2020 · 被引用 26 次
相关 Paper
- Simulation and Retargeting of Complex Multi-Character InteractionsYunbo Zhang, Deepak Gopinath, Yuting Ye, Jessica K. Hodgins 等SIGGRAPH 2023 · 被引用 25 次
- Learning to Sit: Synthesizing Human-Chair Interactions via Hierarchical ControlYu-Wei Chao, Jimei Yang, Weifeng Chen, Jia DengAAAI 2021 · 被引用 50 次
- Zero-Shot Whole-Body Humanoid Control via Behavioral Foundation ModelsAndrea Tirinzoni, Ahmed Touati, Jesse Farebrother, Mateusz Guzek 等ICLR 2025
- Toward Robust Long Range Policy TransferWei-Cheng Tseng, Jin-Siang Lin, Yao-Min Feng, Min SunAAAI 2021 · 被引用 8 次
- InterPrior: Scaling Generative Control for Physics-Based Human-Object InteractionsSirui Xu, Samuel Schulter, Morteza Ziyadi, Xialin He 等CVPR 2026 · 被引用 14 次
