On Provable Benefits of Depth in Training Graph Convolutional Networks
Weilin Cong, Morteza Ramezani, Mehrdad Mahdavi
摘要
Graph Convolutional Networks (GCNs) are known to suffer from performance degradation as the number of layers increases, which is usually attributed to over-smoothing. Despite the apparent consensus, we observe that there exists a discrepancy between the theoretical understanding of over-smoothing and the practical capabilities of GCNs. Specifically, we argue that over-smoothing does not necessarily happen in practice, a deeper model is provably expressive, can converge to global optimum with linear convergence rate, and achieve very high training accuracy as long as properly trained. Despite being capable of achieving high training accuracy, empirical results show that the deeper models generalize poorly on the testing stage and existing theoretical understanding of such behavior remains elusive. To achieve better understanding, we carefully analyze the generalization capability of GCNs, and show that the training strategies to achieve high training accuracy significantly deteriorate the generalization capability of GCNs. Motivated by these findings, we propose a decoupled structure for GCNs that detaches weight matrices from feature propagation to preserve the expressive power and ensure good generalization performance. We conduct empirical evaluations on various synthetic and real-world datasets to validate the correctness of our theory.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper31
- TFE-GNN: A Temporal Fusion Encoder Using Graph Neural Networks for Fine-grained Encrypted Traffic ClassificationHaozhen Zhang, Le Yu, Xi Xiao, Qing Li 等WWW 2023 · 被引用 122 次
- GFT: Graph Foundation Model with Transferable Tree VocabularyZehong Wang, Zheyuan Zhang, Nitesh V. Chawla, Chuxu Zhang 等NeurIPS 2024 · 被引用 108 次
- Demystifying Structural Disparity in Graph Neural Networks: Can One Size Fit All?Haitao Mao, Zhikai Chen, Wei Jin, Haoyu Han 等NeurIPS 2023 · 被引用 58 次
- Model Degradation Hinders Deep Graph Neural NetworksWentao Zhang, Zeang Sheng, Ziqi Yin, Yuezihan Jiang 等KDD 2022 · 被引用 42 次
- Graph Convolutional Kernel Machine versus Graph Convolutional NetworksZhihao Wu, Zhao Zhang, Jicong FanNeurIPS 2023 · 被引用 41 次
它引用的顶会 Paper19
- Open Graph Benchmark: Datasets for Machine Learning on GraphsWeihua Hu, Matthias Fey, Marinka Zitnik, Yuxiao Dong 等NeurIPS 2020 · 被引用 3,935 次
- Simple and Deep Graph Convolutional NetworksMing Chen, Zhewei Wei, Zengfeng Huang, Bolin Ding 等ICML 2020 · 被引用 1,910 次
- DropEdge: Towards Deep Graph Convolutional Networks on Node ClassificationYu Rong, Wenbing Huang, Tingyang Xu, Junzhou HuangICLR 2020 · 被引用 1,599 次
- DeepGCNs: Can GCNs Go As Deep As CNNs?Guohao Li, Matthias Müller, Ali K. Thabet, Bernard GhanemICCV 2019 · 被引用 1,586 次
- Graph Neural Networks Exponentially Lose Expressive Power for Node ClassificationKenta Oono, Taiji SuzukiICLR 2020 · 被引用 864 次
相关 Paper
- Graph Neural Networks Do Not Always OversmoothBastian Epping, Alexandre René, Moritz Helias, Michael T. SchaubNeurIPS 2024 · 被引用 22 次
- Optimization of Graph Neural Networks: Implicit Acceleration by Skip Connections and More DepthKeyulu Xu, Mozhi Zhang, Stefanie Jegelka, Kenji KawaguchiICML 2021 · 被引用 87 次
- DeGNN: Improving Graph Neural Networks with Graph DecompositionXupeng Miao, Nezihe Merve Gürel, Wentao Zhang, Zhichao Han 等KDD 2021 · 被引用 22 次
- Towards Deepening Graph Neural Networks: A GNTK-based Optimization PerspectiveWei Huang, Yayong Li, Weitao Du, Richard Y. D. Xu 等ICLR 2022 · 被引用 19 次
- Dissecting the Diffusion Process in Linear Graph Convolutional NetworksYifei Wang, Yisen Wang, Jiansheng Yang, Zhouchen LinNeurIPS 2021 · 被引用 99 次
