Are we measuring oversmoothing in graph neural networks correctly?
Kaicheng Zhang, Piero Deidda, Desmond Higham, Francesco Tudisco
Abstract
Oversmoothing is a fundamental challenge in graph neural networks (GNNs): as the number of layers increases, node embeddings become increasingly similar, and model performance drops sharply. Traditionally, oversmoothing has been quantified using metrics that measure the similarity of neighbouring node features, such as the Dirichlet energy. We argue that these metrics have critical limitations and fail to reliably capture oversmoothing in realistic scenarios. For instance, they provide meaningful insights only for very deep networks, while typical GNNs show a performance drop already with as few as 10 layers. As an alternative, we propose measuring oversmoothing by examining the numerical or effective rank of the feature representations. We provide extensive numerical evaluation across diverse graph architectures and datasets to show that rank-based metrics consistently capture oversmoothing, whereas energy-based metrics often fail. Notably, we reveal that drops in the rank align closely with performance degradation, even in scenarios where energy metrics remain unchanged. Along with the experimental evaluation, we provide theoretical support for this approach, clarifying why Dirichlet-like measures may fail to capture performance drop and proving that the numerical rank of feature representations collapses to one for a broad family of GNN architectures.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext d199f625-4e9a-458b-8e93-1259e7ef79d7Cited by top-tier papers2
- Sheaf Neural Networks on SPD Manifolds: Second-Order Geometric Representation LearningYuhan Peng, Junwen Dong, Yuzhi Zeng, Hao Li et al.ICML 2026 · 1 citation
- Adaptive Initial Residual Connections for GNNs with Theoretical GuaranteesMohammad Shirzadi, Ali Safarpoor-Dehkordi, Ahad N. ZehmakanAAAI 2026
Builds on23
- How Attentive are Graph Attention Networks?Shaked Brody, Uri Alon, Eran YahavICLR 2022 · 1,717 citations
- DropEdge: Towards Deep Graph Convolutional Networks on Node ClassificationYu Rong, Wenbing Huang, Tingyang Xu, Junzhou HuangICLR 2020 · 1,599 citations
- Measuring and Relieving the Over-Smoothing Problem for Graph Neural Networks from the Topological ViewDeli Chen, Yankai Lin, Wei Li, Peng Li et al.AAAI 2020 · 1,353 citations
- Graph Neural Networks Exponentially Lose Expressive Power for Node ClassificationKenta Oono, Taiji SuzukiICLR 2020 · 864 citations
- PairNorm: Tackling Oversmoothing in GNNsLingxiao Zhao, Leman AkogluICLR 2020 · 590 citations
Related papers
- Dirichlet Energy Constrained Learning for Deep Graph Neural NetworksKaixiong Zhou, Xiao Huang, Daochen Zha, Rui Chen et al.NeurIPS 2021 · 171 citations
- Feature Overcorrelation in Deep Graph Neural Networks: A New PerspectiveWei Jin, Xiaorui Liu, Yao Ma, Charu C. Aggarwal et al.KDD 2022 · 28 citations
- Backward Oversmoothing: why is it hard to train deep Graph Neural Networks?Nicolas KerivenICML 2026 · 4 citations
- Towards Deeper Graph Neural Networks with Differentiable Group NormalizationKaixiong Zhou, Xiao Huang, Yuening Li, Daochen Zha et al.NeurIPS 2020 · 248 citations
- Graph Neural Networks Do Not Always OversmoothBastian Epping, Alexandre René, Moritz Helias, Michael T. SchaubNeurIPS 2024 · 22 citations
