Neural Tangent Kernels Motivate Cross-Covariance Graphs in Neural Networks
Shervin Khalafi, Saurabh Sihag, Alejandro Ribeiro
Abstract
Neural tangent kernels (NTKs) provide a theoretical regime to analyze the learning and generalization behavior of over-parametrized neural networks. For a supervised learning task, the association between the eigenvectors of the NTK and given data (a concept referred to as alignment in this paper) can govern the rate of convergence of gradient descent, as well as generalization to unseen data. Building upon this concept and leveraging the structure of NTKs for graph neural networks (GNNs), we theoretically investigate NTKs and alignment, where our analysis reveals that optimizing the alignment translates to optimizing the graph representation or the graph shift operator (GSO) in a GNN. Our results further establish theoretical guarantees on the optimality of the alignment for a two-layer GNN and these guarantees are characterized by the graph shift operator being a function of the cross-covariance between the input and the output data. The theoretical insights drawn from the analysis of NTKs are validated by our experiments focused on a multi-variate time series prediction task for a publicly available dataset. Specifically, they demonstrate that GNN-based learning models that operate on the cross-covariance matrix indeed outperform those that operate on the covariance matrix estimated from only the input data. Neural Tangent Kernels Motivate Cross-Covariance Graphs in Neural Networks
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers1
Ask how each one uses itBuilds on7
- Spectral Temporal Graph Neural Network for Multivariate Time-series ForecastingDefu Cao, Yujing Wang, Juanyong Duan, Ce Zhang et al.NeurIPS 2020 · 841 citations
- On the linearity of large non-linear models: when and why the tangent kernel is constantChaoyue Liu, Libin Zhu, Mikhail BelkinNeurIPS 2020 · 183 citations
- Understanding Double Descent Requires A Fine-Grained Bias-Variance DecompositionBen Adlam, Jeffrey PenningtonNeurIPS 2020 · 111 citations
- Neural Architecture Search on ImageNet in Four GPU Hours: A Theoretically Inspired PerspectiveWuyang Chen, Xinyu Gong, Zhangyang WangICLR 2021 · 51 citations
- Deep Active Learning by Leveraging Training DynamicsHaonan Wang, Wei Huang, Ziwei Wu, Hanghang Tong et al.NeurIPS 2022 · 49 citations
Related papers
- Temporal Graph Neural Tangent Kernel with Graphon-GuaranteedKatherine Tieu, Dongqi Fu, Yada Zhu, Hendrik F. Hamann et al.NeurIPS 2024 · 14 citations
- Graph Neural Tangent Kernel: Convergence on Large GraphsSanjukta Krishnagopal, Luana RuizICML 2023 · 22 citations
- Understanding the Evolution of the Neural Tangent Kernel at the Edge of StabilityKaiqi Jiang, Jeremy Cohen, Yuanzhi LiNeurIPS 2025 · 8 citations
- How Graph Neural Networks Learn: Lessons from Training DynamicsChenxiao Yang, Qitian Wu, David Wipf, Ruoyu Sun et al.ICML 2024 · 2 citations
- Demystifying GNN-to-MLP Knowledge Transfer: Theoretical Grounding and Dual-Stream Distillation MethodZhiyuan Yu, Mingkai Lin, Wenzhong Li, Zhangyue Yin et al.AAAI 2026
