Expected Value of Communication for Planning in Ad Hoc Teamwork
William Macke, Reuth Mirsky, Peter Stone
Abstract
A desirable goal for autonomous agents is to be able to coordinate on the fly with previously unknown teammates. Known as “ad hoc teamwork”, enabling such a capability has been receiving increasing attention in the research community. One of the central challenges in ad hoc teamwork is quickly recognizing the current plans of other agents and planning accordingly. In this paper, we focus on the scenario in which teammates can communicate with one another, but only at a cost. Thus, they must carefully balance plan recognition based on observations vs. that based on communication.
This paper proposes a new metric for evaluating how similar are two policies that a teammate may be following - the Expected Divergence Point (EDP). We then present a novel planning algorithm for ad hoc teamwork, determining which query to ask and planning accordingly. We demonstrate the effectiveness of this algorithm in a range of increasingly general communication in ad hoc teamwork problems.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 5fa760f4-d9cc-4cd6-9bd5-b57b9319f724Cited by top-tier papers4
- Contrastive Identity-Aware Learning for Multi-Agent Value DecompositionShunyu Liu, Yihe Zhou, Jie Song, Tongya Zheng et al.AAAI 2023 · 43 citations
- Goal Recognition as Reinforcement LearningLeonardo Amado, Reuth Mirsky, Felipe MeneguzziAAAI 2022 · 22 citations
- Autonomous Capability Assessment of Sequential Decision-Making Systems in Stochastic SettingsPulkit Verma, Rushang Karia, Siddharth SrivastavaNeurIPS 2023 · 14 citations
- Inverse Attention Agents for Multi-Agent SystemsQian Long, Ruoyan Li, Minglu Zhao, Tao Gao et al.ICLR 2025
Builds on1
Related papers
- PADiff: Predictive and Adaptive Diffusion Policies for Ad Hoc TeamworkHohei Chan, Xinzhi Zhang, Antao Xiang, Weinan Zhang et al.AAAI 2026
- Online Ad Hoc Teamwork under Partial ObservabilityPengjie Gu, Mengchen Zhao, Jianye Hao, Bo AnICLR 2022 · 35 citations
- Ad Hoc Teamwork via Offline Goal-Based Decision TransformersXinzhi Zhang, Hohei Chan, Deheng Ye, Yi Cai et al.ICML 2025
- Towards Open Ad Hoc Teamwork Using Graph-based Policy LearningArrasy Rahman, Niklas Höpner, Filippos Christianos, Stefano V. AlbrechtICML 2021 · 75 citations
- AATEAM: Achieving the Ad Hoc Teamwork by Employing the Attention MechanismShuo Chen, Ewa Andrejczuk, Zhiguang Cao, Jie ZhangAAAI 2020 · 54 citations
