A Transfer Approach Using Graph Neural Networks in Deep Reinforcement Learning
Tianpei Yang, Heng You, Jianye Hao, Yan Zheng, Matthew E. Taylor
摘要
Transfer learning (TL) has shown great potential to improve Reinforcement Learning (RL) efficiency by leveraging prior knowledge in new tasks. However, much of the existing TL research focuses on transferring knowledge between tasks that share the same state-action spaces. Further, transfer from multiple source tasks that have different state-action spaces is more challenging and needs to be solved urgently to improve the generalization and practicality of the method in real-world scenarios. This paper proposes TURRET (Transfer Using gRaph neuRal nETworks), to utilize the generalization capabilities of Graph Neural Networks (GNNs) to facilitate efficient and effective multi-source policy transfer learning in the state-action mismatch setting. TURRET learns a semantic representation by accounting for the intrinsic property of the agent through GNNs, which leads to a unified state embedding space for all tasks. As a result, TURRET achieves more efficient transfer with strong generalization ability between different tasks and can be easily combined with existing Deep RL algorithms. Experimental results show that TURRET significantly outperforms other TL methods on multiple continuous action control tasks, successfully transferring across robots with different state-action spaces.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper5
- Graph Neural Networks with Adaptive ReadoutsDavid Buterez, Jon Paul Janet, Steven J. Kiddle, Dino Oglic 等NeurIPS 2022 · 被引用 81 次
- Learning Cross-Domain Correspondence for Control with Dynamics Cycle-ConsistencyQiang Zhang, Tete Xiao, Alexei A. Efros, Lerrel Pinto 等ICLR 2021 · 被引用 73 次
- Structure-Aware Transformer Policy for Inhomogeneous Multi-Task Reinforcement LearningSunghoon Hong, Deunsol Yoon, Kee-Eung KimICLR 2022 · 被引用 40 次
- An Efficient Transfer Learning Framework for Multiagent Reinforcement LearningTianpei Yang, Weixun Wang, Hongyao Tang, Jianye Hao 等NeurIPS 2021 · 被引用 34 次
- Snowflake: Scaling GNNs to high-dimensional continuous control via parameter freezingCharlie Blake, Vitaly Kurin, Maximilian Igl, Shimon WhitesonNeurIPS 2021 · 被引用 18 次
相关 Paper
- My Body is a Cage: the Role of Morphology in Graph-Based Incompatible ControlVitaly Kurin, Maximilian Igl, Tim Rocktäschel, Wendelin Boehmer 等ICLR 2021 · 被引用 105 次
- Transfer Value Iteration NetworksJunyi Shen, Hankz Hankui Zhuo, Jin Xu, Bin Zhong 等AAAI 2020 · 被引用 7 次
- REPAINT: Knowledge Transfer in Deep Reinforcement LearningYunzhe Tao, Sahika Genc, Jonathan Chung, Tao Sun 等ICML 2021 · 被引用 32 次
- TempLe: Learning Template of Transitions for Sample Efficient Multi-task RLYanchao Sun, Xiangyu Yin, Furong HuangAAAI 2021 · 被引用 17 次
- Adaptive Transfer Learning on Graph Neural NetworksXueting Han, Zhenhuan Huang, Bang An, Jing BaiKDD 2021 · 被引用 30 次
