Graph Convolutional Reinforcement Learning
Jiechuan Jiang, Chen Dun, Tiejun Huang, Zongqing Lu
摘要
Learning to cooperate is crucially important in multi-agent environments. The key is to understand the mutual interplay between agents. However, multi-agent environments are highly dynamic, where agents keep moving and their neighbors change quickly. This makes it hard to learn abstract representations of mutual interplay between agents. To tackle these difficulties, we propose graph convolutional reinforcement learning, where graph convolution adapts to the dynamics of the underlying graph of the multi-agent environment, and relation kernels capture the interplay between agents by their relation representations. Latent features produced by convolutional layers from gradually increased receptive fields are exploited to learn cooperation, and cooperation is further improved by temporal relation regularization for consistency. Empirically, we show that our method substantially outperforms existing methods in a variety of cooperative scenarios.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper53
- Deep Coordination GraphsWendelin Boehmer, Vitaly Kurin, Shimon WhitesonICML 2020 · 被引用 209 次
- MDP Homomorphic Networks: Group Symmetries in Reinforcement LearningElise van der Pol, Daniel E. Worrall, Herke van Hoof, Frans A. Oliehoek 等NeurIPS 2020 · 被引用 203 次
- Learning Individually Inferred Communication for Multi-Agent CooperationZiluo Ding, Tiejun Huang, Zongqing LuNeurIPS 2020 · 被引用 146 次
- Residual Pathway Priors for Soft Equivariance ConstraintsMarc Finzi, Greg Benton, Andrew Gordon WilsonNeurIPS 2021 · 被引用 89 次
- FOP: Factorizing Optimal Joint Policy of Maximum-Entropy Multi-Agent Reinforcement LearningTianhao Zhang, Yueheng Li, Chen Wang, Guangming Xie 等ICML 2021 · 被引用 88 次
相关 Paper
- Multi-Agent Actor-Critic with Hierarchical Graph Attention NetworkHeechang Ryu, Hayong Shin, Jinkyoo ParkAAAI 2020 · 被引用 143 次
- Reward Propagation Using Graph Convolutional NetworksMartin Klissarov, Doina PrecupNeurIPS 2020 · 被引用 28 次
- Progressive Relation Learning for Group Activity RecognitionGuyue Hu, Bo Cui, Yuan He, Shan YuCVPR 2020
- HYGMA: Hypergraph Coordination Networks with Dynamic Grouping for Multi-Agent Reinforcement LearningChiqiang Liu, Dazi LiICML 2025
- Learning Multi-Agent Communication through Structured Attentive ReasoningMurtaza Rangwala, Ryan WilliamsNeurIPS 2020 · 被引用 42 次
