Towards Open Ad Hoc Teamwork Using Graph-based Policy Learning
Arrasy Rahman, Niklas Höpner, Filippos Christianos, Stefano V. Albrecht
摘要
Ad hoc teamwork is the challenging problem of designing an autonomous agent which can adapt quickly to collaborate with teammates without prior coordination mechanisms, including joint training. Prior work in this area has focused on closed teams in which the number of agents is fixed. In this work, we consider open teams by allowing agents with different fixed policies to enter and leave the environment without prior notification. Our solution builds on graph neural networks to learn agent models and joint-action value models under varying team compositions. We contribute a novel action-value computation that integrates the agent model and joint-action value model to produce action-value estimates. We empirically demonstrate that our approach successfully models the effects other agents have on the learner, leading to policies that robustly adapt to dynamic team compositions and significantly outperform several alternative methods.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper14
- Contrastive Identity-Aware Learning for Multi-Agent Value DecompositionShunyu Liu, Yihe Zhou, Jie Song, Tongya Zheng 等AAAI 2023 · 被引用 43 次
- Online Ad Hoc Teamwork under Partial ObservabilityPengjie Gu, Mengchen Zhao, Jianye Hao, Bo AnICLR 2022 · 被引用 35 次
- N-agent Ad Hoc TeamworkCaroline Wang, Arrasy Rahman, Ishan Durugkar, Elad Liebman 等NeurIPS 2024 · 被引用 21 次
- Minimum Coverage Sets for Training Robust Ad Hoc Teamwork AgentsMuhammad Rahman, Jiaxun Cui, Peter StoneAAAI 2024 · 被引用 20 次
- Fast Peer Adaptation with Context-aware ExplorationLong Ma, Yuanfei Wang, Fangwei Zhong, Song-Chun Zhu 等ICML 2024 · 被引用 9 次
它引用的顶会 Paper3
- Graph Convolutional Reinforcement LearningJiechuan Jiang, Chen Dun, Tiejun Huang, Zongqing LuICLR 2020 · 被引用 415 次
- Shared Experience Actor-Critic for Multi-Agent Reinforcement LearningFilippos Christianos, Lukas Schäfer, Stefano V. AlbrechtNeurIPS 2020 · 被引用 238 次
- Deep Coordination GraphsWendelin Boehmer, Vitaly Kurin, Shimon WhitesonICML 2020 · 被引用 209 次
相关 Paper
- Open Ad Hoc Teamwork with Cooperative Game TheoryJianhong Wang, Yang Li, Yuan Zhang, Wei Pan 等ICML 2024 · 被引用 5 次
- AATEAM: Achieving the Ad Hoc Teamwork by Employing the Attention MechanismShuo Chen, Ewa Andrejczuk, Zhiguang Cao, Jie ZhangAAAI 2020 · 被引用 54 次
- PADiff: Predictive and Adaptive Diffusion Policies for Ad Hoc TeamworkHohei Chan, Xinzhi Zhang, Antao Xiang, Weinan Zhang 等AAAI 2026
- HYGMA: Hypergraph Coordination Networks with Dynamic Grouping for Multi-Agent Reinforcement LearningChiqiang Liu, Dazi LiICML 2025
- Back to the Future: Toward a Hybrid Architecture for Ad Hoc TeamworkHasra Dodampegama, Mohan SridharanAAAI 2023 · 被引用 8 次
