Snowflake: Scaling GNNs to high-dimensional continuous control via parameter freezing
Charlie Blake, Vitaly Kurin, Maximilian Igl, Shimon Whiteson
Abstract
Recent research has shown that graph neural networks (GNNs) can learn policies for locomotion control that are as effective as a typical multi-layer perceptron (MLP), with superior transfer and multi-task performance (Wang et al., 2018; Huang et al., 2020). Results have so far been limited to training on small agents, with the performance of GNNs deteriorating rapidly as the number of sensors and actuators grows. A key motivation for the use of GNNs in the supervised learning setting is their applicability to large graphs, but this benefit has not yet been realised for locomotion control. We identify the weakness with a common GNN architecture that causes this poor scaling: overfitting in the MLPs within the network that encode, decode, and propagate messages. To combat this, we introduce Snowflake, a GNN training method for high-dimensional continuous control that freezes parameters in parts of the network that suffer from overfitting. Snowflake significantly boosts the performance of GNNs for locomotion control on large agents, now matching the performance of MLPs, and with superior transfer properties.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 80f2bebe-068e-4ce4-96d8-d335b7ec6c50Cited by top-tier papers5
- MetaMorph: Learning Universal Controllers with TransformersAgrim Gupta, Linxi Fan, Surya Ganguli, Li Fei-FeiICLR 2022 · 130 citations
- LightTS: Lightweight Time Series Classification with Adaptive Ensemble DistillationDavid Campos, Miao Zhang, Bin Yang, Tung Kieu et al.SIGMOD 2023 · 105 citations
- A Transfer Approach Using Graph Neural Networks in Deep Reinforcement LearningTianpei Yang, Heng You, Jianye Hao, Yan Zheng et al.AAAI 2024 · 4 citations
- MeMo: Meaningful, Modular Controllers via Noise InjectionMegan Tjandrasuwita, Jie Xu, Armando Solar-Lezama, Wojciech MatusikNeurIPS 2024 · 1 citation
- A System for Morphology-Task Generalization via Unified Representation and Behavior DistillationHiroki Furuta, Yusuke Iwasawa, Yutaka Matsuo, Shixiang Shane GuICLR 2023
Builds on9
- One Policy to Control Them All: Shared Modular Policies for Agent-Agnostic ControlWenlong Huang, Igor Mordatch, Deepak PathakICML 2020 · 214 citations
- My Body is a Cage: the Role of Morphology in Graph-Based Incompatible ControlVitaly Kurin, Maximilian Igl, Tim Rocktäschel, Wendelin Boehmer et al.ICLR 2021 · 105 citations
- Randomized Entity-wise Factorization for Multi-Agent Reinforcement LearningShariq Iqbal, Christian A. Schröder de Witt, Bei Peng, Wendelin Boehmer et al.ICML 2021 · 84 citations
- Can Q-Learning with Graph Networks Learn a Generalizable Branching Heuristic for a SAT Solver?Vitaly Kurin, Saad Godil, Shimon Whiteson, Bryan CatanzaroNeurIPS 2020 · 77 citations
- Hard-Coded Gaussian Attention for Neural Machine TranslationWeiqiu You, Simeng Sun, Mohit IyyerACL 2020 · 55 citations
Related papers
- Neural Snowflakes: Universal Latent Graph Inference via Trainable Latent GeometriesHaitz Sáez de Ocáriz Borde, Anastasis KratsiosICLR 2024 · 6 citations
- The Snowflake Hypothesis: Training and Powering GNN with One Node One Receptive FieldKun Wang, Guohao Li, Shilong Wang, Guibin Zhang et al.KDD 2024 · 6 citations
- Graph Neural Networks are Inherently Good Generalizers: Insights by Bridging GNNs and MLPsChenxiao Yang, Qitian Wu, Jiahua Wang, Junchi YanICLR 2023 · 16 citations
- The Heterophilic Snowflake Hypothesis: Training and Empowering GNNs for Heterophilic GraphsKun Wang, Guibin Zhang, Xinnan Zhang, Junfeng Fang et al.KDD 2024 · 6 citations
- Learning to Schedule Learning rate with Graph Neural NetworksYuanhao Xiong, Li-Cheng Lan, Xiangning Chen, Ruochen Wang et al.ICLR 2022 · 18 citations
