Lazy Agents: A New Perspective on Solving Sparse Reward Problem in Multi-agent Reinforcement Learning
Boyin Liu, Zhiqiang Pu, Yi Pan, Jianqiang Yi, Yanyan Liang, Du Zhang
摘要
Sparse reward remains a valuable and challenging problem in multi-agent reinforcement learning (MARL). This paper addresses this issue from a new perspective, i.e., lazy agents. We empirically illustrate how lazy agents damage learning from both exploration and exploitation. Then, we propose a novel MARL framework called Lazy Agents Avoidance through Influencing External States (LAIES). Firstly, we examine the causes and types of lazy agents in MARL using a causal graph of the interaction between agents and their environment. Then, we mathematically define the concept of fully lazy agents and teams by calculating the causal effect of their actions on external states using the do-calculus process. Based on definitions, we provide two intrinsic rewards to motivate agents, i.e., individual diligence intrinsic motivation (IDI) and collaborative diligence intrinsic motivation (CDI). IDI and CDI employ counterfactual reasoning based on the external states transition model (ESTM) we developed. Empirical results demonstrate that our proposed method achieves state-of-the-art performance on various tasks, including the sparsereward version of StarCraft multi-agent challenge (SMAC) and Google Research Football (GRF). Our code is open-source and available at https: //github.com/liuboyin/LAIES .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper12
- LLM-Empowered State Representation for Reinforcement LearningBoyuan Wang, Yun Qu, Yuhang Jiang, Jianzhun Shao 等ICML 2024 · 被引用 34 次
- FoX: Formation-Aware Exploration in Multi-Agent Reinforcement LearningYonghyeon Jo, Sunwoo Lee, Junghyuk Yeom, Seungyul HanAAAI 2024 · 被引用 22 次
- Unlocking the Power of Multi-Agent LLM for Reasoning: From Lazy Agents to DeliberationZhiwei Zhang, Xiaomin Li, Yudi Lin, Hui Liu 等ICLR 2026 · 被引用 13 次
- Understanding Individual Agent Importance in Multi-Agent System via Counterfactual ReasoningJianming Chen, Yawen Wang, Junjie Wang, Xiaofei Xie 等AAAI 2025 · 被引用 11 次
- Individual Contributions as Intrinsic Exploration Scaffolds for Multi-agent Reinforcement LearningXinran Li, Zifan Liu, Shibo Chen, Jun ZhangICML 2024 · 被引用 11 次
它引用的顶会 Paper7
- Weighted QMIX: Expanding Monotonic Value Function Factorisation for Deep Multi-Agent Reinforcement LearningTabish Rashid, Gregory Farquhar, Bei Peng, Shimon WhitesonNeurIPS 2020 · 被引用 1,960 次
- Google Research Football: A Novel Reinforcement Learning EnvironmentKarol Kurach, Anton Raichuk, Piotr Stanczyk, Michal Zajac 等AAAI 2020 · 被引用 496 次
- Celebrating Diversity in Shared Multi-Agent Reinforcement LearningChenghao Li, Tonghan Wang, Chengjie Wu, Qianchuan Zhao 等NeurIPS 2021 · 被引用 224 次
- Cooperative Exploration for Multi-Agent Deep Reinforcement LearningIou-Jen Liu, Unnat Jain, Raymond A. Yeh, Alexander G. SchwingICML 2021 · 被引用 133 次
- PMIC: Improving Multi-Agent Reinforcement Learning with Progressive Mutual Information CollaborationPengyi Li, Hongyao Tang, Tianpei Yang, Xiaotian Hao 等ICML 2022 · 被引用 49 次
相关 Paper
- Causality-Aware Efficient Exploration for Cooperative Multi-Agent Reinforcement LearningHongye Cao, Tianpei Yang, Fan Feng, Hammadi Rafik Ouariachi 等AAAI 2026
- LIGS: Learnable Intrinsic-Reward Generation Selection for Multi-Agent LearningDavid Henry Mguni, Taher Jafferjee, Jianhong Wang, Nicolas Perez Nieves 等ICLR 2022 · 被引用 20 次
- TMAE: Learning Targeted Multi-Agent Exploration via Causal InferenceChuxiong Sun, Dunqi Yao, Rui Wang, Wenwen Qiang 等AAAI 2026
- Subspace-Aware Exploration for Sparse-Reward Multi-Agent TasksPei Xu, Junge Zhang, Qiyue Yin, Chao Yu 等AAAI 2023 · 被引用 15 次
- ELIGN: Expectation Alignment as a Multi-Agent Intrinsic RewardZixian Ma, Rose E. Wang, Fei-Fei Li, Michael S. Bernstein 等NeurIPS 2022 · 被引用 22 次
