Enhancing Human Experience in Human-Agent Collaboration: A Human-Centered Modeling Approach Based on Positive Human Gain
Yiming Gao, Feiyu Liu, Liang Wang, Dehua Zheng, Zhenjie Lian, Weixuan Wang, Wenjin Yang, Siqin Li, Xianliang Wang, Wenhui Chen, Jing Dai, Qiang Fu
Abstract
Existing game AI research mainly focuses on enhancing agents' abilities to win games, but this does not inherently make humans have a better experience when collaborating with these agents. For example, agents may dominate the collaboration and exhibit unintended or detrimental behaviors, leading to poor experiences for their human partners. In other words, most game AI agents are modeled in a "self-centered" manner. In this paper, we propose a "human-centered" modeling scheme for collaborative agents that aims to enhance the experience of humans. Specifically, we model the experience of humans as the goals they expect to achieve during the task. We expect that agents should learn to enhance the extent to which humans achieve these goals while maintaining agents' original abilities (e.g., winning games). To achieve this, we propose the Reinforcement Learning from Human Gain (RLHG) approach. The RLHG approach introduces a "baseline", which corresponds to the extent to which humans primitively achieve their goals, and encourages agents to learn behaviors that can effectively enhance humans in achieving their goals better. We evaluate the RLHG agent in the popular Multi-player Online Battle Arena (MOBA) game, Honor of Kings, by conducting real-world human-agent tests. Both objective performance and subjective preference results show that the RLHG agent provides participants better gaming experience. INTRODUCTION Recently, Reinforcement Learning (RL) has been widely used in developing Artificial Intelligence (AI) systems for games, developing various agents that perform at a human-level performance, such as AlphaGo (Silver et al., 2016; 2017 ) in Go, AlphaStar (Vinyals et al., 2019) in StarCraftII, OpenAI Five (OpenAI et al., 2019) in Dota2, and Wukong AI (Ye et al., 2020a) in Honor of Kings. To further expand the applications of these agents, researchers are exploring ways to improve their generalization to human partners (
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers2
- Exploring Collaboration Mechanisms for LLM Agents: A Social Psychology ViewJintian Zhang, Xin Xu, Ningyu Zhang, Ruibo Liu et al.ACL 2024
- Characterizing an LLM-driven Social Network: The Case of Chirper.aiYiming Zhu, Yupeng He, Ehsan-Ul Haq, Gareth Tyson et al.CSCW 2026
Builds on10
- Training language models to follow instructions with human feedbackLong Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida et al.NeurIPS 2022 · 24,707 citations
- Mastering Complex Control in MOBA Games with Deep Reinforcement LearningDeheng Ye, Zhao Liu, Mingfei Sun, Bei Shi et al.AAAI 2020 · 395 citations
- "Other-Play" for Zero-Shot CoordinationHengyuan Hu, Adam Lerer, Alex Peysakhovich, Jakob N. FoersterICML 2020 · 271 citations
- Collaborating with Humans without Human DataDJ Strouse, Kevin R. McKee, Matt M. Botvinick, Edward Hughes et al.NeurIPS 2021 · 239 citations
- Towards Playing Full MOBA Games with Deep Reinforcement LearningDeheng Ye, Guibin Chen, Wen Zhang, Sheng Chen et al.NeurIPS 2020 · 225 citations
Related papers
- Towards Effective and Interpretable Human-Agent Collaboration in MOBA Games: A Communication PerspectiveYiming Gao, Feiyu Liu, Liang Wang, Zhenjie Lian et al.ICLR 2023 · 2 citations
- Learning Diverse Policies in MOBA Games via Macro-GoalsYiming Gao, Bei Shi, Xueying Du, Liang Wang et al.NeurIPS 2021 · 17 citations
- Online-to-Offline RL for Agent AlignmentXu Liu, Haobo Fu, Stefano V. Albrecht, Qiang Fu et al.ICLR 2025
- Learning to Incentivize Other Learning AgentsJiachen Yang, Ang Li, Mehrdad Farajtabar, Peter Sunehag et al.NeurIPS 2020 · 105 citations
- Rating-Based Reinforcement LearningDevin White, Mingkang Wu, Ellen R. Novoseller, Vernon J. Lawhern et al.AAAI 2024 · 10 citations
