Theory of Mind for Multi-Agent Collaboration via Large Language Models
Huao Li, Yu Quan Chong, Simon Stepputtis, Joseph Campbell, Dana Hughes, Charles Lewis, Katia P. Sycara
Abstract
While Large Language Models (LLMs) have demonstrated impressive accomplishments in both reasoning and planning, their abilities in multi-agent collaborations remains largely unexplored. This study evaluates LLM-based agents in a multi-agent cooperative text game with Theory of Mind (ToM) inference tasks, comparing their performance with Multi-Agent Reinforcement Learning (MARL) and planning-based baselines. We observed evidence of emergent collaborative behaviors and high-order Theory of Mind capabilities among LLM-based agents. Our results reveal limitations in LLM-based agents' planning optimization due to systematic failures in managing long-horizon contexts and hallucination about the task state. We explore the use of explicit belief state representations to mitigate these issues, finding that it enhances task performance and the accuracy of ToM inferences for LLM-based agents.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext cc0a83fb-9a78-4e29-a63e-7ff2af3a7d9dCited by top-tier papers33
- Multi-LLM Debate: Framework, Principals, and InterventionsAndrew Estornell, Yang LiuNeurIPS 2024 · 131 citations
- CoRAL: Collaborative Retrieval-Augmented Large Language Models Improve Long-tail RecommendationJunda Wu, Cheng-Chun Chang, Tong Yu, Zhankui He et al.KDD 2024 · 32 citations
- Language Grounded Multi-agent Reinforcement Learning with Human-interpretable CommunicationHuao Li, Hossein Nourkhiz Mahjoub, Behdad Chalaki, Vaishnav Tadiparthi et al.NeurIPS 2024 · 31 citations
- Shapley-Coop: Credit Assignment for Emergent Cooperation in Self-Interested LLM AgentsYun Hua, Haosheng Chen, Shiqin Wang, Wenhao Li et al.NeurIPS 2025 · 13 citations
- TOM-SWE: User Mental Modeling For Software Engineering AgentsXuhui Zhou, Valerie Chen, Zhiruo Wang, Graham Neubig et al.ICML 2026 · 12 citations
Builds on5
- Chain-of-Thought Prompting Elicits Reasoning in Large Language ModelsJason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma et al.NeurIPS 2022 · 22,562 citations
- Generative Agents: Interactive Simulacra of Human BehaviorJoon Sung Park, Joseph C. O'Brien, Carrie Jun Cai, Meredith Ringel Morris et al.UIST 2023 · 1,882 citations
- Language Models as Zero-Shot Planners: Extracting Actionable Knowledge for Embodied AgentsWenlong Huang, Pieter Abbeel, Deepak Pathak, Igor MordatchICML 2022 · 1,539 citations
- Neural Theory-of-Mind? On the Limits of Social Intelligence in Large LMsMaarten Sap, Ronan Le Bras, Daniel Fried, Yejin ChoiEMNLP 2022 · 92 citations
- Minding Language Models' (Lack of) Theory of Mind: A Plug-and-Play Multi-Character Belief TrackerMelanie Sclar, Sachin Kumar, Peter West, Alane Suhr et al.ACL 2023 · 21 citations
Related papers
- Adaptive Theory of Mind for LLM-based Multi-Agent CoordinationChunjiang Mu, Ya Zeng, Qiaosheng Zhang, Kun Shao et al.AAAI 2026
- Language Models Represent Beliefs of Self and OthersWentao Zhu, Zhining Zhang, Yizhou WangICML 2024 · 24 citations
- Hypothetical Minds: Scaffolding Theory of Mind for Multi-Agent Tasks with Large Language ModelsLogan Matthew Cross, Violet Xiang, Agam Bhatia, Daniel L. K. Yamins et al.ICLR 2025 · 2 citations
- MetaMind: Modeling Human Social Thoughts with Metacognitive Multi-Agent SystemsXuanming Zhang, Yuxuan Chen, Samuel (Min-Hsuan) Yeh, Sharon LiNeurIPS 2025 · 14 citations
- AutoToM: Scaling Model-based Mental Inference via Automated Agent ModelingZhining Zhang, Chuanyang Jin, Mung Yao Jia, Shunchi Zhang et al.NeurIPS 2025 · 30 citations
