Learning to Brachiate via Simplified Model Imitation
Daniele Reda, Hung Yu Ling, Michiel van de Panne
Abstract
Brachiation is the primary form of locomotion for gibbons and siamangs, in which these primates swing from tree limb to tree limb using only their arms. It is challenging to control because of the limited control authority, the required advance planning, and the precision of the required grasps. We present a novel approach to this problem using reinforcement learning, and as demonstrated on a finger-less 14-link planar model that learns to brachiate across challenging handhold sequences. Key to our method is the use of a simplified model, a point mass with a virtual arm, for which we first learn a policy that can brachiate across handhold sequences with a prescribed order. This facilitates the learning of the policy for the full model, for which it provides guidance by providing an overall center-of-mass trajectory to imitate, as well as for the timing of the holds. Lastly, the simplified model can also readily be used for planning suitable sequences of handholds in a given environment. Our results demonstrate brachiation motions with a variety of durations for the flight and hold phases, as well as emergent extra back-and-forth swings when this proves useful. The system is evaluated with a variety of ablations. The method enables future work towards more general 3D brachiation, as well as using simplified model imitation in other settings. For videos, supplementary material and code, visit: https://brachiation-rl.github.io/brachiation.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 8a33d462-2e71-4dde-a4e9-05f296acaa7dCited by top-tier papers1
Ask how each one uses itBuilds on3
- A scalable approach to control diverse behaviors for physically simulated charactersJungdam Won, Deepak Gopinath, Jessica K. HodginsSIGGRAPH 2020 · 146 citations
- Fast and flexible multilegged locomotion using learned centroidal dynamicsTaesoo Kwon, Yoonsang Lee, Michiel van de PanneSIGGRAPH 2020 · 50 citations
- CARL: controllable agent with reinforcement learning for quadruped locomotionYing-Sheng Luo, Jonathan Hans Soeseno, Trista Pei-Chun Chen, Wei-Chao ChenSIGGRAPH 2020 · 49 citations
Related papers
- D-Grasp: Physically Plausible Dynamic Grasp Synthesis for Hand-Object InteractionsSammy Joe Christen, Muhammed Kocabas, Emre Aksan, Jemin Hwangbo et al.CVPR 2022 · 69 citations
- Omnigrasp: Grasping Diverse Objects with Simulated HumanoidsZhengyi Luo, Jinkun Cao, Sammy Christen, Alexander Winkler et al.NeurIPS 2024 · 66 citations
- Breathing Life Into Biomechanical User ModelsAleksi Ikkala, Florian Fischer, Markus Klar, Miroslav Bachinski et al.UIST 2022 · 33 citations
- Predicting Mid-Air Interaction Movements and Fatigue Using Deep Reinforcement LearningNoshaba Cheema, Laura A. Frey-Law, Kourosh Naderi, Jaakko Lehtinen et al.CHI 2020 · 66 citations
- Hand-Object Interaction Controller (HOIC): Deep Reinforcement Learning for Reconstructing Interactions with PhysicsHaoyu Hu, Xinyu Yi, Zhe Cao, Jun-Hai Yong et al.SIGGRAPH 2024 · 2 citations
