Structure-Aware Transformer Policy for Inhomogeneous Multi-Task Reinforcement Learning
Sunghoon Hong, Deunsol Yoon, Kee-Eung Kim
摘要
Modular Reinforcement Learning, where the agent is assumed to be morphologically structured as a graph, for example composed of limbs and joints, aims to learn a policy that is transferable to a structurally similar but different agent. Compared to traditional Multi-Task Reinforcement Learning, this promising approach allows us to cope with inhomogeneous tasks where the state and action space dimensions differ across tasks. Graph Neural Networks are a natural model for representing the pertinent policies, but a recent work has shown that their multi-hop message passing mechanism is not ideal for conveying important information to other modules and thus a transformer model without morphological information was proposed. In this work, we argue that the morphological information is still very useful and propose a transformer policy model that effectively encodes such information. Specifically, we encode the morphological information in terms of the traversal-based positional embedding and the graph-based relational embedding. We empirically show that the morphological information is crucial for modular reinforcement learning, substantially outperforming prior state-of-the-art methods on multi-task learning as well as transfer learning settings with different state and action space dimensions.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
引用它的顶会 Paper17
- Universal Morphology Control via Contextual ModulationZheng Xiong, Jacob Beck, Shimon WhitesonICML 2023 · 被引用 27 次
- Low-Rank Modular Reinforcement Learning via Muscle SynergyHeng Dong, Tonghan Wang, Jiayuan Liu, Chongjie ZhangNeurIPS 2022 · 被引用 21 次
- Subequivariant Graph Reinforcement Learning in 3D EnvironmentsRunfa Chen, Jiaqi Han, Fuchun Sun, Wenbing HuangICML 2023 · 被引用 14 次
- Centralized Reward Agent for Knowledge Sharing and Transfer in Multi-Task Reinforcement LearningHaozhe Ma, Zhengding Luo, Thanh Vinh Vo, Kuankuan Sima 等NeurIPS 2025 · 被引用 9 次
- Distilling Morphology-Conditioned Hypernetworks for Efficient Universal Morphology ControlZheng Xiong, Risto Vuorio, Jacob Beck, Matthieu Zimmer 等ICML 2024 · 被引用 8 次
相关 Paper
- My Body is a Cage: the Role of Morphology in Graph-Based Incompatible ControlVitaly Kurin, Maximilian Igl, Tim Rocktäschel, Wendelin Boehmer 等ICLR 2021 · 被引用 105 次
- One Policy to Control Them All: Shared Modular Policies for Agent-Agnostic ControlWenlong Huang, Igor Mordatch, Deepak PathakICML 2020 · 被引用 214 次
- MetaMorph: Learning Universal Controllers with TransformersAgrim Gupta, Linxi Fan, Surya Ganguli, Li Fei-FeiICLR 2022 · 被引用 130 次
- MeMo: Meaningful, Modular Controllers via Noise InjectionMegan Tjandrasuwita, Jie Xu, Armando Solar-Lezama, Wojciech MatusikNeurIPS 2024 · 被引用 1 次
- A Transfer Approach Using Graph Neural Networks in Deep Reinforcement LearningTianpei Yang, Heng You, Jianye Hao, Yan Zheng 等AAAI 2024 · 被引用 4 次
