Neurosymbolic Transformers for Multi-Agent Communication
Jeevana Priya Inala, Yichen Yang, James Paulos, Yewen Pu, Osbert Bastani, Vijay Kumar, Martin C. Rinard, Armando Solar-Lezama
摘要
We study the problem of inferring communication structures that can solve cooperative multi-agent planning problems while minimizing the amount of communication. We quantify the amount of communication as the maximum degree of the communication graph; this metric captures settings where agents have limited bandwidth. Minimizing communication is challenging due to the combinatorial nature of both the decision space and the objective; for instance, we cannot solve this problem by training neural networks using gradient descent. We propose a novel algorithm that synthesizes a control policy that combines a programmatic communication policy used to generate the communication graph with a transformer policy network used to choose actions. Our algorithm first trains the transformer policy, which implicitly generates a "soft" communication graph; then, it synthesizes a programmatic communication policy that "hardens" this graph, forming a neurosymbolic transformer. Our experiments demonstrate how our approach can synthesize policies that generate low-degree communication graphs while maintaining near-optimal performance.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper11
- Cooperative Exploration for Multi-Agent Deep Reinforcement LearningIou-Jen Liu, Unnat Jain, Raymond A. Yeh, Alexander G. SchwingICML 2021 · 被引用 133 次
- Compositional Reinforcement Learning from Logical SpecificationsKishor Jothimurugan, Suguman Bansal, Osbert Bastani, Rajeev AlurNeurIPS 2021 · 被引用 112 次
- Learning Transformer ProgramsDan Friedman, Alexander Wettig, Danqi ChenNeurIPS 2023 · 被引用 59 次
- Web question answering with neurosymbolic program synthesisQiaochu Chen, Aaron Lamoreaux, Xinyu Wang, Greg Durrett 等PLDI 2021 · 被引用 25 次
- Delayed Propagation Transformer: A Universal Computation Engine towards Practical Control in Cyber-Physical SystemsWenqing Zheng, Qiangqiang Guo, Hao Yang, Peihao Wang 等NeurIPS 2021 · 被引用 13 次
它引用的顶会 Paper1
相关 Paper
- Bridging Training and Execution via Dynamic Directed Graph-Based Communication in Cooperative Multi-Agent SystemsZhuohui Zhang, Bin He, Bin Cheng, Gang LiAAAI 2025 · 被引用 10 次
- What Planning Problems Can A Relational Neural Network Solve?Jiayuan Mao, Tomás Lozano-Pérez, Joshua B. Tenenbaum, Leslie Pack KaelblingNeurIPS 2023 · 被引用 13 次
- Learning Multi-Agent Communication from Graph Modeling PerspectiveShengchao Hu, Li Shen, Ya Zhang, Dacheng TaoICLR 2024 · 被引用 65 次
- Transform2Act: Learning a Transform-and-Control Policy for Efficient Agent DesignYe Yuan, Yuda Song, Zhengyi Luo, Wen Sun 等ICLR 2022 · 被引用 51 次
- Discovering symbolic policies with deep reinforcement learningMikel Landajuela, Brenden K. Petersen, Sookyung Kim, Cláudio P. Santiago 等ICML 2021 · 被引用 118 次
