Scaling Multi-Agent Reinforcement Learning with Selective Parameter Sharing
Filippos Christianos, Georgios Papoudakis, Arrasy Rahman, Stefano V. Albrecht
摘要
Sharing parameters in multi-agent deep reinforcement learning has played an essential role in allowing algorithms to scale to a large number of agents. Parameter sharing between agents significantly decreases the number of trainable parameters, shortening training times to tractable levels, and has been linked to more efficient learning. However, having all agents share the same parameters can also have a detrimental effect on learning. We demonstrate the impact of parameter sharing methods on training speed and converged returns, establishing that when applied indiscriminately, their effectiveness is highly dependent on the environment. We propose a novel method to automatically identify agents which may benefit from sharing parameters by partitioning them based on their abilities and goals. Our approach combines the increased sample efficiency of parameter sharing with the representational capacity of multiple independent networks to reduce training time and increase final returns.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper33
- Celebrating Diversity in Shared Multi-Agent Reinforcement LearningChenghao Li, Tonghan Wang, Chengjie Wu, Qianchuan Zhao 等NeurIPS 2021 · 被引用 224 次
- Towards a Standardised Performance Evaluation Protocol for Cooperative MARLRihab Gorsane, Omayma Mahjoub, Ruan de Kock, Roland Dubb 等NeurIPS 2022 · 被引用 79 次
- Efficient Multi-agent Communication via Self-supervised Information AggregationCong Guan, Feng Chen, Lei Yuan, Chenghe Wang 等NeurIPS 2022 · 被引用 65 次
- LDSA: Learning Dynamic Subtask Assignment in Cooperative Multi-Agent Reinforcement LearningMingyu Yang, Jian Zhao, Xunhan Hu, Wengang Zhou 等NeurIPS 2022 · 被引用 61 次
- CooHOI: Learning Cooperative Human-Object Interaction with Manipulated Object DynamicsJiawei Gao, Ziqin Wang, Zeqi Xiao, Jingbo Wang 等NeurIPS 2024 · 被引用 57 次
它引用的顶会 Paper4
- ROMA: Multi-Agent Reinforcement Learning with Emergent RolesTonghan Wang, Heng Dong, Victor R. Lesser, Chongjie ZhangICML 2020 · 被引用 286 次
- Shared Experience Actor-Critic for Multi-Agent Reinforcement LearningFilippos Christianos, Lukas Schäfer, Stefano V. AlbrechtNeurIPS 2020 · 被引用 238 次
- Succinct and Robust Multi-Agent Communication With Temporal Message ControlSai Qian Zhang, Qi Zhang, Jieyu LinNeurIPS 2020 · 被引用 90 次
- Learning Multi-Agent Communication through Structured Attentive ReasoningMurtaza Rangwala, Ryan WilliamsNeurIPS 2020 · 被引用 42 次
相关 Paper
- Kaleidoscope: Learnable Masks for Heterogeneous Multi-agent Reinforcement LearningXinran Li, Ling Pan, Jun ZhangNeurIPS 2024 · 被引用 10 次
- HyperMARL: Adaptive Hypernetworks for Multi-Agent RLKale-ab Abebe Tessera, Arrasy Rahman, Amos J. Storkey, Stefano V. AlbrechtNeurIPS 2025 · 被引用 11 次
- Towards Complete Multi-Agent Coordination Policy Learning via Denoising Maximum Entropy OptimizationGuanghao Li, lei yuan, Ruiqi Xue, Hengchang Zhang 等ICML 2026
- Reward Dimension Reduction for Scalable Multi-Objective Reinforcement LearningGiseung Park, Youngchul SungICLR 2025
- Sharing Knowledge in Multi-Task Deep Reinforcement LearningCarlo D'Eramo, Davide Tateo, Andrea Bonarini, Marcello Restelli 等ICLR 2020 · 被引用 148 次
