Dynamic population-based meta-learning for multi-agent communication with natural language
Abhinav Gupta, Marc Lanctot, Angeliki Lazaridou
Abstract
In this work, our goal is to train agents that can coordinate with seen, unseen as well as human partners in a multi-agent communication environment involving natural language. Previous work using a single set of agents has shown great progress in generalizing to known partners, however it struggles when coordinating with unfamiliar agents. To mitigate that, recent work explored the use of population-based approaches, where multiple agents interact with each other with the goal of learning more generic protocols. These methods, while able to result in good coordination between unseen partners, still only achieve so in cases of simple languages, thus failing to adapt to human partners using natural language. We attribute this to the use of static populations and instead propose a dynamic population-based meta-learning approach that builds such a population in an iterative manner. We perform a holistic evaluation of our method on two different referential games, and show that our agents outperform all prior work when communicating with seen partners and humans. Furthermore, we analyze the natural language generation skills of our agents, where we find that our agents also outperform strong baselines. Finally, we test the robustness of our agents when communicating with out-of-population agents and carefully test the importance of each component of our method through ablation studies.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext fdd5e1d0-d50c-443a-8db7-6dbc1b2928b1Cited by top-tier papers5
- Multi-Agent Meta-Reinforcement Learning: Sharper Convergence Rates with Task SimilarityWeichao Mao, Haoran Qiu, Chen Wang, Hubertus Franke et al.NeurIPS 2023 · 17 citations
- Learning Multi-Agent Communication with Contrastive LearningYat Long Lo, Biswa Sengupta, Jakob Nicolaus Foerster, Michael NoukhovitchICLR 2024 · 11 citations
- Learning Zero-Shot Cooperation with Humans, Assuming Humans Are BiasedChao Yu, Jiaxuan Gao, Weilin Liu, Botian Xu et al.ICLR 2023 · 4 citations
- Learning Multi-Object Positional Relationships via Emergent CommunicationYicheng Feng, Boshi An, Zongqing LuAAAI 2024 · 4 citations
- Towards Effective and Interpretable Human-Agent Collaboration in MOBA Games: A Communication PerspectiveYiming Gao, Feiyu Liu, Liang Wang, Zhenjie Lian et al.ICLR 2023 · 2 citations
Builds on5
- "Other-Play" for Zero-Shot CoordinationHengyuan Hu, Adam Lerer, Alex Peysakhovich, Jakob N. FoersterICML 2020 · 271 citations
- Countering Language Drift with Seeded Iterated LearningYuchen Lu, Soumye Singhal, Florian Strub, Aaron C. Courville et al.ICML 2020 · 85 citations
- Compositionality and Generalization In Emergent LanguagesRahma Chaabouni, Eugene Kharitonov, Diane Bouchacourt, Emmanuel Dupoux et al.ACL 2020 · 40 citations
- On the interaction between supervision and self-play in emergent communicationRyan Lowe, Abhinav Gupta, Jakob N. Foerster, Douwe Kiela et al.ICLR 2020 · 30 citations
- Multi-agent Communication meets Natural Language: Synergies between Functional and Structural Language LearningAngeliki Lazaridou, Anna Potapenko, Olivier TielemanACL 2020 · 11 citations
Related papers
- Revisiting Populations in multi-agent CommunicationPaul Michel, Mathieu Rita, Kory Wallace Mathewson, Olivier Tieleman et al.ICLR 2023
- Emergent Communication of GeneralizationsJesse Mu, Noah D. GoodmanNeurIPS 2021 · 60 citations
- Efficient Reinforcement Learning for Zero-Shot Coordination in Evolving GamesBingyu Hui, Lebin Yu, Quanming Yao, Yunpeng Qu et al.AAAI 2026 · 1 citation
- Few-shot Language Coordination by Modeling Theory of MindHao Zhu, Graham Neubig, Yonatan BiskICML 2021 · 43 citations
- Unsupervised Partner Design Enables Robust Ad-hoc TeamworkConstantin Ruhdorfer, Matteo Bortoletto, Victor Oei, Anna Penzkofer et al.ICML 2026 · 3 citations
