Adaptive Initial Residual Connections for GNNs with Theoretical Guarantees
Mohammad Shirzadi, Ali Safarpoor-Dehkordi, Ahad N. Zehmakan
Abstract
Message passing is the core operation in graph neural networks, where each node updates its embeddings by aggregating information from its neighbors. However, in deep architectures, this process often leads to diminished expressiveness. A popular solution is the use of residual connections, where the input from the current (or initial) layer is added to the aggregated neighbor information to preserve embeddings across layers. Following a recent line of research, we investigate an adaptive residual scheme in which different nodes have varying residual strengths. We prove that this approach prevents oversmoothing; particularly, we show that the Dirichlet energy of the embeddings remains bounded away from zero. This is the first theoretical guarantee not only for the adaptive setting, but also for static residual connections (where residual strengths are shared across nodes) with activation functions. Furthermore, based on an extensive set of experiments, this adaptive approach is shown to outperform the standard and state-of-the-art message passing mechanisms, especially on heterophilic graphs. To improve the time complexity of our approach, we introduce a variant in which residual strengths are not learned but instead set heuristically, a choice that performs as well as the learnable version.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 7b41127c-a4c3-4d5c-aaa1-9ea6b663895fCited by top-tier papers1
Ask how each one uses itBuilds on22
- Simple and Deep Graph Convolutional NetworksMing Chen, Zhewei Wei, Zengfeng Huang, Bolin Ding et al.ICML 2020 · 1,910 citations
- DeepGCNs: Can GCNs Go As Deep As CNNs?Guohao Li, Matthias Müller, Ali K. Thabet, Bernard GhanemICCV 2019 · 1,586 citations
- Measuring and Relieving the Over-Smoothing Problem for Graph Neural Networks from the Topological ViewDeli Chen, Yankai Lin, Wei Li, Peng Li et al.AAAI 2020 · 1,353 citations
- Recipe for a General, Powerful, Scalable Graph TransformerLadislav Rampásek, Michael Galkin, Vijay Prakash Dwivedi, Anh Tuan Luu et al.NeurIPS 2022 · 1,216 citations
- Graph Neural Networks Exponentially Lose Expressive Power for Node ClassificationKenta Oono, Taiji SuzukiICLR 2020 · 864 citations
Related papers
- Residual Connections and Normalization Can Provably Prevent Oversmoothing in GNNsMichael Scholkemper, Xinyi Wu, Ali Jadbabaie, Michael T. SchaubICLR 2025
- Difference Residual Graph Neural NetworksLiang Yang, Weihang Peng, Wenmiao Zhou, Bingxin Niu et al.ACM MM 2022 · 6 citations
- Graph Neural Networks with Adaptive ResidualXiaorui Liu, Jiayuan Ding, Wei Jin, Han Xu et al.NeurIPS 2021 · 100 citations
- A Signed Graph Approach to Understanding and Mitigating OversmoothingJiaqi Wang, Xinyi Wu, James Cheng, Yifei WangNeurIPS 2025 · 4 citations
- Mitigating Oversmoothing Through Reverse Process of GNNs for Heterophilic GraphsMoonjeong Park, Jaeseung Heo, Dongwoo KimICML 2024 · 7 citations
