Reward Propagation Using Graph Convolutional Networks
Martin Klissarov, Doina Precup
摘要
Potential-based reward shaping provides an approach for designing good reward functions, with the purpose of speeding up learning. However, automatically finding potential functions for complex environments is a difficult problem (in fact, of the same difficulty as learning a value function from scratch). We propose a new framework for learning potential functions by leveraging ideas from graph representation learning. Our approach relies on Graph Convolutional Networks which we use as a key ingredient in combination with the probabilistic inference view of reinforcement learning. More precisely, we leverage Graph Convolutional Networks to perform message passing from rewarding states. The propagated messages can then be used as potential functions for reward shaping to accelerate learning. We verify empirically that our approach can achieve considerable improvements in both small and high-dimensional control problems.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- Flexible Option LearningMartin Klissarov, Doina PrecupNeurIPS 2021 · 被引用 38 次
- Deep Laplacian-based Options for Temporally-Extended ExplorationMartin Klissarov, Marlos C. MachadoICML 2023 · 被引用 31 次
- Neural Algorithmic Reasoners are Implicit PlannersAndreea Deac, Petar Velickovic, Ognjen Milinkovic, Pierre-Luc Bacon 等NeurIPS 2021 · 被引用 27 次
- Snowflake: Scaling GNNs to high-dimensional continuous control via parameter freezingCharlie Blake, Vitaly Kurin, Maximilian Igl, Shimon WhitesonNeurIPS 2021 · 被引用 18 次
相关 Paper
- Automatic Reward Shaping from Confounded Offline DataMingxuan Li, Junzhe Zhang, Elias BareinboimICML 2025
- Towards Better Laplacian Representation in Reinforcement Learning with Generalized Graph DrawingKaixin Wang, Kuangqi Zhou, Qixin Zhang, Jie Shao 等ICML 2021 · 被引用 32 次
- Graph Convolutional Reinforcement LearningJiechuan Jiang, Chen Dun, Tiejun Huang, Zongqing LuICLR 2020 · 被引用 415 次
- Reachability-Aware Laplacian Representation in Reinforcement LearningKaixin Wang, Kuangqi Zhou, Jiashi Feng, Bryan Hooi 等ICML 2023 · 被引用 10 次
- Graph Reinforcement Learning for Network Control via Bi-Level OptimizationDaniele Gammelli, James Harrison, Kaidi Yang, Marco Pavone 等ICML 2023 · 被引用 14 次
