Learning to Brachiate via Simplified Model Imitation
Daniele Reda, Hung Yu Ling, Michiel van de Panne
摘要
Brachiation is the primary form of locomotion for gibbons and siamangs, in which these primates swing from tree limb to tree limb using only their arms. It is challenging to control because of the limited control authority, the required advance planning, and the precision of the required grasps. We present a novel approach to this problem using reinforcement learning, and as demonstrated on a finger-less 14-link planar model that learns to brachiate across challenging handhold sequences. Key to our method is the use of a simplified model, a point mass with a virtual arm, for which we first learn a policy that can brachiate across handhold sequences with a prescribed order. This facilitates the learning of the policy for the full model, for which it provides guidance by providing an overall center-of-mass trajectory to imitate, as well as for the timing of the holds. Lastly, the simplified model can also readily be used for planning suitable sequences of handholds in a given environment. Our results demonstrate brachiation motions with a variety of durations for the flight and hold phases, as well as emergent extra back-and-forth swings when this proves useful. The system is evaluated with a variety of ablations. The method enables future work towards more general 3D brachiation, as well as using simplified model imitation in other settings. For videos, supplementary material and code, visit: https://brachiation-rl.github.io/brachiation.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper3
- A scalable approach to control diverse behaviors for physically simulated charactersJungdam Won, Deepak Gopinath, Jessica K. HodginsSIGGRAPH 2020 · 被引用 146 次
- Fast and flexible multilegged locomotion using learned centroidal dynamicsTaesoo Kwon, Yoonsang Lee, Michiel van de PanneSIGGRAPH 2020 · 被引用 50 次
- CARL: controllable agent with reinforcement learning for quadruped locomotionYing-Sheng Luo, Jonathan Hans Soeseno, Trista Pei-Chun Chen, Wei-Chao ChenSIGGRAPH 2020 · 被引用 49 次
相关 Paper
- D-Grasp: Physically Plausible Dynamic Grasp Synthesis for Hand-Object InteractionsSammy Joe Christen, Muhammed Kocabas, Emre Aksan, Jemin Hwangbo 等CVPR 2022 · 被引用 69 次
- Omnigrasp: Grasping Diverse Objects with Simulated HumanoidsZhengyi Luo, Jinkun Cao, Sammy Christen, Alexander Winkler 等NeurIPS 2024 · 被引用 66 次
- Breathing Life Into Biomechanical User ModelsAleksi Ikkala, Florian Fischer, Markus Klar, Miroslav Bachinski 等UIST 2022 · 被引用 33 次
- Predicting Mid-Air Interaction Movements and Fatigue Using Deep Reinforcement LearningNoshaba Cheema, Laura A. Frey-Law, Kourosh Naderi, Jaakko Lehtinen 等CHI 2020 · 被引用 66 次
- Hand-Object Interaction Controller (HOIC): Deep Reinforcement Learning for Reconstructing Interactions with PhysicsHaoyu Hu, Xinyu Yi, Zhe Cao, Jun-Hai Yong 等SIGGRAPH 2024 · 被引用 2 次
