Language Grounded Multi-agent Reinforcement Learning with Human-interpretable Communication
Huao Li, Hossein Nourkhiz Mahjoub, Behdad Chalaki, Vaishnav Tadiparthi, Kwonjoon Lee, Ehsan Moradi-Pari, Charles Lewis, Katia P. Sycara
摘要
Multi-Agent Reinforcement Learning (MARL) methods have shown promise in enabling agents to learn a shared communication protocol from scratch and accomplish challenging team tasks. However, the learned language is usually not interpretable to humans or other agents not co-trained together, limiting its applicability in ad-hoc teamwork scenarios. In this work, we propose a novel computational pipeline that aligns the communication space between MARL agents with an embedding space of human natural language by grounding agent communications on synthetic data generated by embodied Large Language Models (LLMs) in interactive teamwork scenarios. Our results demonstrate that introducing language grounding not only maintains task performance but also accelerates the emergence of communication. Furthermore, the learned communication protocols exhibit zero-shot generalization capabilities in ad-hoc teamwork scenarios with unseen teammates and novel task states. This work presents a significant step toward enabling effective communication and collaboration between artificial agents and humans in real-world teamwork settings.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- Adaptive Collaboration with Humans: Metacognitive Policy Optimization for Multi-Agent LLMs with Continual LearningWei Yang, Defu Cao, Jiacheng Pang, Muyan Weng 等ICLR 2026 · 被引用 11 次
- Weaving in the Clouds: Achieving Synergistic Collaboration among LLM Agents via Federated LearningJiaxing Zhao, Hongbin Xie, Yuzhen Lei, Xuan Song 等ICML 2026
- TACTIC: Task-Aware Sparse Coordination Graphs for Multi-Task Multi-agent Reinforcement LearningKexing Peng, Pengyi Li, tinghuai ma, Jianye HaoICML 2026
- Learning Efficient and Interpretable Multi-Agent CommunicationWei Du, Benyu Wu, Yuqing Sun, Wei Guo 等ICLR 2026
- M2I2: Learning Efficient Multi-Agent Communication via Masked State Modeling and Intention InferenceChuxiong Sun, Peng He, Qirui Ji, Zehua Zang 等AAAI 2026
它引用的顶会 Paper15
- Generative Agents: Interactive Simulacra of Human BehaviorJoon Sung Park, Joseph C. O'Brien, Carrie Jun Cai, Meredith Ringel Morris 等UIST 2023 · 被引用 1,882 次
- Grounding Large Language Models in Interactive Environments with Online Reinforcement LearningThomas Carta, Clément Romac, Thomas Wolf, Sylvain Lamprier 等ICML 2023 · 被引用 258 次
- Language Models Meet World Models: Embodied Experiences Enhance Language ModelsJiannan Xiang, Tianhua Tao, Yi Gu, Tianmin Shu 等NeurIPS 2023 · 被引用 180 次
- Learning to Ground Multi-Agent Communication with AutoencodersToru Lin, Jacob Huh, Christopher Stauffer, Ser-Nam Lim 等NeurIPS 2021 · 被引用 75 次
- Emergent Communication at ScaleRahma Chaabouni, Florian Strub, Florent Altché, Eugene Tarassov 等ICLR 2022 · 被引用 65 次
相关 Paper
- LLM-Assisted Semantically Diverse Teammate Generation for Efficient Multi-agent CoordinationLihe Li, Lei Yuan, Pengsen Liu, Tao Jiang 等ICML 2025
- Emergent Discrete Communication in Semantic SpacesMycal Tucker, Huao Li, Siddharth Agrawal, Dana Hughes 等NeurIPS 2021 · 被引用 34 次
- ProAgent: Building Proactive Cooperative Agents with Large Language ModelsCeyao Zhang, Kaijie Yang, Siyi Hu, Zihao Wang 等AAAI 2024 · 被引用 141 次
- Theory of Mind for Multi-Agent Collaboration via Large Language ModelsHuao Li, Yu Quan Chong, Simon Stepputtis, Joseph Campbell 等EMNLP 2023 · 被引用 57 次
- LLM-Planner: Few-Shot Grounded Planning for Embodied Agents with Large Language ModelsChan Hee Song, Brian M. Sadler, Jiaman Wu, Wei-Lun Chao 等ICCV 2023 · 被引用 685 次
