LIGS: Learnable Intrinsic-Reward Generation Selection for Multi-Agent Learning
David Henry Mguni, Taher Jafferjee, Jianhong Wang, Nicolas Perez Nieves, Oliver Slumbers, Feifei Tong, Yang Li, Jiangcheng Zhu, Yaodong Yang, Jun Wang
Abstract
Efficient exploration is important for reinforcement learners to achieve high rewards. In multi-agent systems, coordinated exploration and behaviour is critical for agents to jointly achieve optimal outcomes. In this paper, we introduce a new general framework for improving coordination and performance of multi-agent reinforcement learners (MARL). Our framework, named Learnable Intrinsic-Reward Generation Selection algorithm (LIGS) introduces an adaptive learner, Generator that observes the agents and learns to construct intrinsic rewards online that coordinate the agents’ joint exploration and joint behaviour. Using a novel combination of MARL and switching controls, LIGS determines the best states to learn to add intrinsic rewards which leads to a highly efficient learning process. LIGS can subdivide complex tasks making them easier to solve and enables systems of MARL agents to quickly solve environments with sparse rewards. LIGS can seamlessly adopt existing MARL algorithms and, our theory shows that it ensures convergence to policies that deliver higher system performance. We demonstrate its superior performance in challenging tasks in Foraging and StarCraft II.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext ec37bbf4-698f-4752-bd71-81ac1b3eb37aCited by top-tier papers5
- Multi-Agent Reinforcement Learning is a Sequence Modeling ProblemMuning Wen, Jakub Grudzien Kuba, Runji Lin, Weinan Zhang et al.NeurIPS 2022 · 408 citations
- Towards a Standardised Performance Evaluation Protocol for Cooperative MARLRihab Gorsane, Omayma Mahjoub, Ruan de Kock, Roland Dubb et al.NeurIPS 2022 · 79 citations
- Intrinsic Action Tendency Consistency for Cooperative Multi-Agent Reinforcement LearningJunkai Zhang, Yifan Zhang, Xi Sheryl Zhang, Yifan Zang et al.AAAI 2024 · 9 citations
- Open Ad Hoc Teamwork with Cooperative Game TheoryJianhong Wang, Yang Li, Yuan Zhang, Wei Pan et al.ICML 2024 · 5 citations
- Vision-Based Generic Potential Function for Policy Alignment in Multi-Agent Reinforcement LearningHao Ma, Shijie Wang, Zhiqiang Pu, Siyao Zhao et al.AAAI 2025 · 1 citation
Builds on6
- Multi-Agent Reinforcement Learning for Active Voltage Control on Power Distribution NetworksJianhong Wang, Wangkun Xu, Yunjie Gu, Wenbin Song et al.NeurIPS 2021 · 216 citations
- Shapley Q-Value: A Local Reward Approach to Solve Global Reward GamesJianhong Wang, Yuan Zhang, Tae-Kyun Kim, Yunjie GuAAAI 2020 · 159 citations
- Settling the Variance of Multi-Agent Policy GradientsJakub Grudzien Kuba, Muning Wen, Linghui Meng, Shangding Gu et al.NeurIPS 2021 · 121 citations
- Multi-Agent Determinantal Q-LearningYaodong Yang, Ying Wen, Jun Wang, Liheng Chen et al.ICML 2020 · 83 citations
- Learning in Nonzero-Sum Stochastic Games with PotentialsDavid Henry Mguni, Yutong Wu, Yali Du, Yaodong Yang et al.ICML 2021 · 51 citations
Related papers
- MASER: Multi-Agent Reinforcement Learning with Subgoals Generated from Experience Replay BufferJeewon Jeon, Woojun Kim, Whiyoung Jung, Youngchul SungICML 2022 · 53 citations
- MANSA: Learning Fast and Slow in Multi-Agent SystemsDavid Henry Mguni, Haojun Chen, Taher Jafferjee, Jianhong Wang et al.ICML 2023 · 4 citations
- Individual Contributions as Intrinsic Exploration Scaffolds for Multi-agent Reinforcement LearningXinran Li, Zifan Liu, Shibo Chen, Jun ZhangICML 2024 · 11 citations
- Lazy Agents: A New Perspective on Solving Sparse Reward Problem in Multi-agent Reinforcement LearningBoyin Liu, Zhiqiang Pu, Yi Pan, Jianqiang Yi et al.ICML 2023 · 34 citations
- Generative Exploration and ExploitationJiechuan Jiang, Zongqing LuAAAI 2020 · 6 citations
