AATEAM: Achieving the Ad Hoc Teamwork by Employing the Attention Mechanism
Shuo Chen, Ewa Andrejczuk, Zhiguang Cao, Jie Zhang
摘要
In the ad hoc teamwork setting, a team of agents needs to perform a task without prior coordination. The most advanced approach learns policies based on previous experiences and reuses one of the policies to interact with new teammates. However, the selected policy in many cases is sub-optimal. Switching between policies to adapt to new teammates' behaviour takes time, which threatens the successful performance of a task. In this paper, we propose AATEAM – a method that uses the attention-based neural networks to cope with new teammates' behaviour in real-time. We train one attention network per teammate type. The attention networks learn both to extract the temporal correlations from the sequence of states (i.e. contexts) and the mapping from contexts to actions. Each attention network also learns to predict a future state given the current context and its output action. The prediction accuracies help to determine which actions the ad hoc agent should take. We perform extensive experiments to show the effectiveness of our method.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper8
- Online Ad Hoc Teamwork under Partial ObservabilityPengjie Gu, Mengchen Zhao, Jianye Hao, Bo AnICLR 2022 · 被引用 35 次
- Very Important Person Localization in Unconstrained Conditions: A New BenchmarkXiao Wang, Zheng Wang, Toshihiko Yamasaki, Wenjun ZengAAAI 2021 · 被引用 11 次
- Back to the Future: Toward a Hybrid Architecture for Ad Hoc TeamworkHasra Dodampegama, Mohan SridharanAAAI 2023 · 被引用 8 次
- Open Ad Hoc Teamwork with Cooperative Game TheoryJianhong Wang, Yang Li, Yuan Zhang, Wei Pan 等ICML 2024 · 被引用 5 次
- Controlling Type Confounding in Ad Hoc Teamwork with Instance-wise Teammate Feedback RectificationDong Xing, Pengjie Gu, Qian Zheng, Xinrun Wang 等ICML 2023 · 被引用 4 次
相关 Paper
- Towards Open Ad Hoc Teamwork Using Graph-based Policy LearningArrasy Rahman, Niklas Höpner, Filippos Christianos, Stefano V. AlbrechtICML 2021 · 被引用 75 次
- PADiff: Predictive and Adaptive Diffusion Policies for Ad Hoc TeamworkHohei Chan, Xinzhi Zhang, Antao Xiang, Weinan Zhang 等AAAI 2026
- Ad Hoc Teamwork via Offline Goal-Based Decision TransformersXinzhi Zhang, Hohei Chan, Deheng Ye, Yi Cai 等ICML 2025
- Learning Multi-Agent Communication through Structured Attentive ReasoningMurtaza Rangwala, Ryan WilliamsNeurIPS 2020 · 被引用 42 次
- Multi-Agent Actor-Critic with Hierarchical Graph Attention NetworkHeechang Ryu, Hayong Shin, Jinkyoo ParkAAAI 2020 · 被引用 143 次
