DEP-RL: Embodied Exploration for Reinforcement Learning in Overactuated and Musculoskeletal Systems
Pierre Schumacher, Daniel F. B. Haeufle, Dieter Büchler, Syn Schmitt, Georg Martius
摘要
Muscle-actuated organisms are capable of learning an unparalleled diversity of dexterous movements despite their vast amount of muscles. Reinforcement learning (RL) on large musculoskeletal models, however, has not been able to show similar performance. We conjecture that ineffective exploration in large overactuated action spaces is a key problem. This is supported by our finding that common exploration noise strategies are inadequate in synthetic examples of overactuated systems. We identify differential extrinsic plasticity (DEP), a method from the domain of self-organization, as being able to induce state-space covering exploration within seconds of interaction. By integrating DEP into RL, we achieve fast learning of reaching and locomotion in musculoskeletal systems, outperforming current approaches in all considered tasks in sample efficiency and robustness. 1
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper13
- Latent exploration for Reinforcement LearningAlberto Silvio Chiappa, Alessandro Marin Vargas, Ann Zixiang Huang, Alexander MathisNeurIPS 2023 · 被引用 41 次
- MyoDex: A Generalizable Prior for Dexterous ManipulationVittorio Caggiano, Sudeep Dasari, Vikash KumarICML 2023 · 被引用 26 次
- DynSyn: Dynamical Synergistic Representation for Efficient Learning and Control in Overactuated Embodied SystemsKaibo He, Chenhui Zuo, Chengtian Ma, Yanan SuiICML 2024 · 被引用 19 次
- Open the Black Box: Step-based Policy Updates for Temporally-Correlated Episodic Reinforcement LearningGe Li, Hongyi Zhou, Dominik Roth, Serge Thilges 等ICLR 2024 · 被引用 11 次
- SIM2VR: Towards Automated Biomechanical Testing in VRFlorian Fischer, Aleksi Ikkala, Markus Klar, Arthur Fleig 等UIST 2024 · 被引用 10 次
它引用的顶会 Paper4
- Parrot: Data-Driven Behavioral Priors for Reinforcement LearningAvi Singh, Huihan Liu, Gaoyue Zhou, Albert Yu 等ICLR 2021 · 被引用 161 次
- Growing Action SpacesGregory Farquhar, Laura Gustafson, Zeming Lin, Shimon Whiteson 等ICML 2020 · 被引用 48 次
- When should agents explore?Miruna Pislar, David Szepesvari, Georg Ostrovski, Diana L. Borsa 等ICLR 2022 · 被引用 26 次
- Learning to Represent Action Values as a Hypergraph on the Action VerticesArash Tavakoli, Mehdi Fatemi, Petar KormushevICLR 2021 · 被引用 25 次
相关 Paper
- Explore to Learn: Latent Exploration Through Disentangled Synergy Patterns for Reinforcement Learning in Overactuated ControlYiming Wang, Kaiyan Zhao, Xu Li, Yan Li 等AAAI 2026 · 被引用 1 次
- Accelerated Policy Learning with Parallel Differentiable SimulationJie Xu, Viktor Makoviychuk, Yashraj Narang, Fabio Ramos 等ICLR 2022 · 被引用 141 次
- Simplified Temporal Consistency Reinforcement LearningYi Zhao, Wenshuai Zhao, Rinu Boney, Juho Kannala 等ICML 2023 · 被引用 19 次
- MetaCURE: Meta Reinforcement Learning with Empowerment-Driven ExplorationJin Zhang, Jianhao Wang, Hao Hu, Tong Chen 等ICML 2021 · 被引用 33 次
- Low-Rank Modular Reinforcement Learning via Muscle SynergyHeng Dong, Tonghan Wang, Jiayuan Liu, Chongjie ZhangNeurIPS 2022 · 被引用 21 次
