Towards Deep Attention in Graph Neural Networks: Problems and Remedies
Soo Yong Lee, Fanchen Bu, Jaemin Yoo, Kijung Shin
摘要
Graph neural networks (GNNs) learn the representation of graph-structured data, and their expressiveness can be further enhanced by inferring node relations for propagation. Attention-based GNNs infer neighbor importance to manipulate the weight of its propagation. Despite their popularity, the discussion on deep graph attention and its unique challenges has been limited. In this work, we investigate some problematic phenomena related to deep graph attention, including vulnerability to over-smoothed features and smooth cumulative attention. Through theoretical and empirical analyses, we show that various attention-based GNNs suffer from these problems. Motivated by our findings, we propose AERO-GNN, a novel GNN architecture designed for deep graph attention. AERO-GNN provably mitigates the proposed problems of deep graph attention, which is further empirically demonstrated with (a) its adaptive and less smooth attention functions and (b) higher performance at deep layers (up to 64). On 9 out of 12 node classification benchmarks, AERO-GNN outperforms the baseline GNNs, highlighting the advantages of deep graph attention. Our code is available at https: //github.com/syleeheal/AERO-GNN .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper13
- Rethinking Reconstruction-based Graph-Level Anomaly Detection: Limitations and a Simple RemedySunwoo Kim, Soo Yong Lee, Fanchen Bu, Shinhwan Kang 等NeurIPS 2024 · 被引用 27 次
- On Which Nodes Does GCN Fail? Enhancing GCN From the Node PerspectiveJincheng Huang, Jialie Shen, Xiaoshuang Shi, Xiaofeng ZhuICML 2024 · 被引用 19 次
- How Interpretable Are Interpretable Graph Neural Networks?Yongqiang Chen, Yatao Bian, Bo Han, James ChengICML 2024 · 被引用 17 次
- Sign is Not a Remedy: Multiset-to-Multiset Message Passing for Learning on Heterophilic GraphsLangzhang Liang, Sunwoo Kim, Kijung Shin, Zenglin Xu 等ICML 2024 · 被引用 13 次
- Feature Distribution on Graph Topology Mediates the Effect of Graph Convolution: Homophily PerspectiveSoo Yong Lee, Sunwoo Kim, Fanchen Bu, Jaemin Yoo 等ICML 2024 · 被引用 9 次
它引用的顶会 Paper36
- Simple and Deep Graph Convolutional NetworksMing Chen, Zhewei Wei, Zengfeng Huang, Bolin Ding 等ICML 2020 · 被引用 1,910 次
- How Attentive are Graph Attention Networks?Shaked Brody, Uri Alon, Eran YahavICLR 2022 · 被引用 1,717 次
- DeepGCNs: Can GCNs Go As Deep As CNNs?Guohao Li, Matthias Müller, Ali K. Thabet, Bernard GhanemICCV 2019 · 被引用 1,586 次
- Geom-GCN: Geometric Graph Convolutional NetworksHongbin Pei, Bingzhe Wei, Kevin Chen-Chuan Chang, Yu Lei 等ICLR 2020 · 被引用 1,445 次
- Learning to Simulate Complex Physics with Graph NetworksAlvaro Sanchez-Gonzalez, Jonathan Godwin, Tobias Pfaff, Rex Ying 等ICML 2020 · 被引用 1,439 次
相关 Paper
- Towards Deeper Graph Neural NetworksMeng Liu, Hongyang Gao, Shuiwang JiKDD 2020 · 被引用 496 次
- Dirichlet Energy Constrained Learning for Deep Graph Neural NetworksKaixiong Zhou, Xiao Huang, Daochen Zha, Rui Chen 等NeurIPS 2021 · 被引用 171 次
- Improving Breadth-Wise Backpropagation in Graph Neural Networks Helps Learning Long-Range DependenciesDenis Lukovnikov, Asja FischerICML 2021 · 被引用 16 次
- Deep Attention Diffusion Graph Neural Networks for Text ClassificationYonghao Liu, Renchu Guan, Fausto Giunchiglia, Yanchun Liang 等EMNLP 2021 · 被引用 67 次
- Enhancing Graph Representations Learning with Decorrelated PropagationHua Liu, Haoyu Han, Wei Jin, Xiaorui Liu 等KDD 2023 · 被引用 7 次
