Controlling Type Confounding in Ad Hoc Teamwork with Instance-wise Teammate Feedback Rectification
Dong Xing, Pengjie Gu, Qian Zheng, Xinrun Wang, Shanqi Liu, Longtao Zheng, Bo An, Gang Pan
Abstract
Ad hoc teamwork requires an agent to cooperate with unknown teammates without prior coordination. Many works propose to abstract teammate instances into high-level representation of types and then pre-train the best response for each type. However, most of them do not consider the distribution of teammate instances within a type. This could expose the agent to the hidden risk of type confounding. In the worst case, the best response for an abstract teammate type could be the worst response for all specific instances of that type. This work addresses the issue from the lens of causal inference. We first theoretically demonstrate that this phenomenon is due to the spurious correlation brought by uncontrolled teammate distribution. Then, we propose our solution, CTCAT, which disentangles such correlation through an instance-wise teammate feedback rectification. This operation reweights the interaction of teammate instances within a shared type to reduce the influence of type confounding. The effect of CTCAT is evaluated in multiple domains, including classic ad hoc teamwork tasks and real-world scenarios. Results show that CTCAT is robust to the influence of type confounding, a practical issue that directly hazards the robustness of our trained agents but was unnoticed in previous works.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 1f21349a-3c99-4bab-b40a-5208fc16e753Builds on9
- An Optimistic Perspective on Offline Reinforcement LearningRishabh Agarwal, Dale Schuurmans, Mohammad NorouziICML 2020 · 568 citations
- Learning to Reach Goals via Iterated Supervised LearningDibya Ghosh, Abhishek Gupta, Ashwin Reddy, Justin Fu et al.ICLR 2021 · 222 citations
- Agent Modelling under Partial Observability for Deep Reinforcement LearningGeorgios Papoudakis, Filippos Christianos, Stefano V. AlbrechtNeurIPS 2021 · 110 citations
- Towards Open Ad Hoc Teamwork Using Graph-based Policy LearningArrasy Rahman, Niklas Höpner, Filippos Christianos, Stefano V. AlbrechtICML 2021 · 75 citations
- AATEAM: Achieving the Ad Hoc Teamwork by Employing the Attention MechanismShuo Chen, Ewa Andrejczuk, Zhiguang Cao, Jie ZhangAAAI 2020 · 54 citations
Related papers
- Knowing Unknown Teammates: Exploring Anonymity and Explanations in a Teammate Information-Sharing Recommender SystemGeoff Musick, Elizabeth S. Gilman, Wen Duan, Nathan J. McNeese et al.CSCW 2023 · 4 citations
- Online Ad Hoc Teamwork under Partial ObservabilityPengjie Gu, Mengchen Zhao, Jianye Hao, Bo AnICLR 2022 · 35 citations
- Back to the Future: Toward a Hybrid Architecture for Ad Hoc TeamworkHasra Dodampegama, Mohan SridharanAAAI 2023 · 8 citations
- Ad Hoc Teamwork via Offline Goal-Based Decision TransformersXinzhi Zhang, Hohei Chan, Deheng Ye, Yi Cai et al.ICML 2025
- PADiff: Predictive and Adaptive Diffusion Policies for Ad Hoc TeamworkHohei Chan, Xinzhi Zhang, Antao Xiang, Weinan Zhang et al.AAAI 2026
