Model Degradation Hinders Deep Graph Neural Networks
Wentao Zhang, Zeang Sheng, Ziqi Yin, Yuezihan Jiang, Yikuan Xia, Jun Gao, Zhi Yang, Bin Cui
摘要
Graph Neural Networks (GNNs) have achieved great success in various graph mining tasks. However, drastic performance degradation is always observed when a GNN is stacked with many layers. As a result, most GNNs only have shallow architectures, which limits their expressive power and exploitation of deep neighborhoods. Most recent studies attribute the performance degradation of deep GNNs to the over-smoothing issue. In this paper, we disentangle the conventional graph convolution operation into two independent operations: Propagation (P) and Transformation (T). Following this, the depth of a GNN can be split into the propagation depth (𝐷 𝑝 ) and the transformation depth (𝐷 𝑡 ). Through extensive experiments, we find that the major cause for the performance degradation of deep GNNs is the model degradation issue caused by large 𝐷 𝑡 rather than the over-smoothing issue mainly caused by large 𝐷 𝑝 . Further, we present Adaptive Initial Residual (AIR), a plug-and-play module compatible with all kinds of GNN architectures, to alleviate the model degradation issue and the over-smoothing issue simultaneously. Experimental results on six real-world datasets demonstrate that GNNs equipped with AIR outperform most GNNs with shallow architectures owing to the benefits of both large 𝐷 𝑝 and 𝐷 𝑡 , while the time costs associated with AIR can be ignored. CCS CONCEPTS • Computing methodologies → Machine learning; • Mathematics of computing → Graph algorithms.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper14
- Rethinking Propagation for Unsupervised Graph Domain AdaptationMeihan Liu, Zeyu Fang, Zhen Zhang, Ming Gu 等AAAI 2024 · 被引用 45 次
- Towards Deep Attention in Graph Neural Networks: Problems and RemediesSoo Yong Lee, Fanchen Bu, Jaemin Yoo, Kijung ShinICML 2023 · 被引用 44 次
- Learning Strong Graph Neural Networks with Weak InformationYixin Liu, Kaize Ding, Jianling Wang, Vincent C. S. Lee 等KDD 2023 · 被引用 40 次
- Oversmoothing, "Oversquashing", Heterophily, Long-Range, and more: Demystifying Common Beliefs in Graph Machine LearningAdrián Arnaiz-Rodríguez, Federico ErricaICLR 2026 · 被引用 26 次
- Graph-Skeleton: 1% Nodes are Sufficient to Represent Billion-Scale GraphLinfeng Cao, Haoran Deng, Yang Yang, Chunping Wang 等WWW 2024 · 被引用 15 次
它引用的顶会 Paper22
- LightGCN: Simplifying and Powering Graph Convolution Network for RecommendationXiangnan He, Kuan Deng, Xiang Wang, Yan Li 等SIGIR 2020 · 被引用 4,448 次
- Open Graph Benchmark: Datasets for Machine Learning on GraphsWeihua Hu, Matthias Fey, Marinka Zitnik, Yuxiao Dong 等NeurIPS 2020 · 被引用 3,935 次
- Simple and Deep Graph Convolutional NetworksMing Chen, Zhewei Wei, Zengfeng Huang, Bolin Ding 等ICML 2020 · 被引用 1,910 次
- DropEdge: Towards Deep Graph Convolutional Networks on Node ClassificationYu Rong, Wenbing Huang, Tingyang Xu, Junzhou HuangICLR 2020 · 被引用 1,599 次
- DeepGCNs: Can GCNs Go As Deep As CNNs?Guohao Li, Matthias Müller, Ali K. Thabet, Bernard GhanemICCV 2019 · 被引用 1,586 次
相关 Paper
- Towards Deeper Graph Neural NetworksMeng Liu, Hongyang Gao, Shuiwang JiKDD 2020 · 被引用 496 次
- DRGCN: Dynamic Evolving Initial Residual for Deep Graph Convolutional NetworksLei Zhang, Xiaodong Yan, Jianshan He, Ruopeng Li 等AAAI 2023 · 被引用 17 次
- DeGNN: Improving Graph Neural Networks with Graph DecompositionXupeng Miao, Nezihe Merve Gürel, Wentao Zhang, Zhichao Han 等KDD 2021 · 被引用 22 次
- Orthogonal Graph Neural NetworksKai Guo, Kaixiong Zhou, Xia Hu, Yu Li 等AAAI 2022 · 被引用 41 次
- Difference Residual Graph Neural NetworksLiang Yang, Weihang Peng, Wenmiao Zhou, Bingxin Niu 等ACM MM 2022 · 被引用 6 次
