The Sensory Neuron as a Transformer: Permutation-Invariant Neural Networks for Reinforcement Learning
Yujin Tang, David Ha
摘要
In complex systems, we often observe complex global behavior emerge from a collection of agents interacting with each other in their environment, with each individual agent acting only on locally available information, without knowing the full picture. Such systems have inspired development of artificial intelligence algorithms in areas such as swarm optimization and cellular automata. Motivated by the emergence of collective behavior from complex cellular systems, we build systems that feed each sensory input from the environment into distinct, but identical neural networks, each with no fixed relationship with one another. We show that these sensory networks can be trained to integrate information received locally, and through communication via an attention mechanism, can collectively produce a globally coherent policy. Moreover, the system can still perform its task even if the ordering of its inputs is randomly permuted several times during an episode. These permutation invariant systems also display useful robustness and generalization properties that are broadly applicable. Interactive demo and videos of our results: https://attentionneuron.github.io/
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper8
- Variational Neural Cellular AutomataRasmus Berg Palm, Miguel González Duque, Shyam Sudhakaran, Sebastian RisiICLR 2022 · 被引用 35 次
- Inducing Point Operator Transformer: A Flexible and Scalable Architecture for Solving PDEsSeungjun Lee, Taeil OhAAAI 2024 · 被引用 22 次
- Discovering Evolution Strategies via Meta-Black-Box OptimizationRobert Tjarko Lange, Tom Schaul, Yutian Chen, Tom Zahavy 等ICLR 2023 · 被引用 21 次
- Relational Reasoning via Set Transformers: Provable Efficiency and Applications to MARLFengzhuo Zhang, Boyi Liu, Kaixin Wang, Vincent Y. F. Tan 等NeurIPS 2022 · 被引用 16 次
- Permutation Equivariance of Transformers and its ApplicationsHengyuan Xu, Liyao Xiang, Hangyu Ye, Dixi Yao 等CVPR 2024 · 被引用 11 次
它引用的顶会 Paper14
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah 等NeurIPS 2020 · 被引用 64,255 次
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- Extracting Training Data from Large Language ModelsNicholas Carlini, Florian Tramèr, Eric Wallace, Matthew Jagielski 等USENIX Security 2021 · 被引用 2,866 次
- Nyströmformer: A Nyström-based Algorithm for Approximating Self-AttentionYunyang Xiong, Zhanpeng Zeng, Rudrasis Chakraborty, Mingxing Tan 等AAAI 2021 · 被引用 675 次
- Recurrent Independent MechanismsAnirudh Goyal, Alex Lamb, Jordan Hoffmann, Shagun Sodhani 等ICLR 2021 · 被引用 357 次
相关 Paper
- Seeing the forest and the tree: Building representations of both individual and collective dynamics with transformersRan Liu, Mehdi Azabou, Max Dabagia, Jingyun Xiao 等NeurIPS 2022 · 被引用 28 次
- Consensus Learning for Cooperative Multi-Agent Reinforcement LearningZhiwei Xu, Bin Zhang, Dapeng Li, Zeren Zhang 等AAAI 2023 · 被引用 27 次
- Multi-Agent MDP Homomorphic NetworksElise van der Pol, Herke van Hoof, Frans A. Oliehoek, Max WellingICLR 2022 · 被引用 36 次
- One Policy to Control Them All: Shared Modular Policies for Agent-Agnostic ControlWenlong Huang, Igor Mordatch, Deepak PathakICML 2020 · 被引用 214 次
- Learning Intuitive Policies Using Action FeaturesMingwei Ma, Jizhou Liu, Samuel Sokota, Max Kleiman-Weiner 等ICML 2023 · 被引用 4 次
