Fisher Information Embedding for Node and Graph Learning
Dexiong Chen, Paolo Pellizzoni, Karsten M. Borgwardt
Abstract
Attention-based graph neural networks (GNNs), such as graph attention networks (GATs), have become popular neural architectures for processing graph-structured data and learning node embeddings. Despite their empirical success, these models rely on labeled data and the theoretical properties of these models have yet to be fully understood. In this work, we propose a novel attention-based node embedding framework for graphs. Our framework builds upon a hierarchical kernel for multisets of subgraphs around nodes (e.g., neighborhoods) and each kernel leverages the geometry of a smooth statistical manifold to compare pairs of multisets, by "projecting" the multisets onto the manifold. By explicitly computing node embeddings with a manifold of Gaussian mixtures, our method leads to a new attention mechanism for neighborhood aggregation. We provide theoretical insights into generalizability and expressivity of our embeddings, contributing to a deeper understanding of attention-based GNNs. We propose both efficient unsupervised and supervised methods for learning the embeddings. Through experiments on several node classification benchmarks, we demonstrate that our proposed method outperforms existing attention-based graph models like GATs. Our code is available at https://github.com/BorgwardtLab/ fisher_information_embedding .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers1
Ask how each one uses itBuilds on11
- Open Graph Benchmark: Datasets for Machine Learning on GraphsWeihua Hu, Matthias Fey, Marinka Zitnik, Yuxiao Dong et al.NeurIPS 2020 · 3,935 citations
- How Attentive are Graph Attention Networks?Shaked Brody, Uri Alon, Eran YahavICLR 2022 · 1,717 citations
- Do Transformers Really Perform Badly for Graph Representation?Chengxuan Ying, Tianle Cai, Shengjie Luo, Shuxin Zheng et al.NeurIPS 2021 · 1,632 citations
- Recipe for a General, Powerful, Scalable Graph TransformerLadislav Rampásek, Michael Galkin, Vijay Prakash Dwivedi, Anh Tuan Luu et al.NeurIPS 2022 · 1,216 citations
- Structure-Aware Transformer for Graph Representation LearningDexiong Chen, Leslie O'Bray, Karsten M. BorgwardtICML 2022 · 349 citations
Related papers
- Multi-View Representation Learning with Manifold SmoothnessShu Li, Wei Wang, Wen-Tao Li, Pan ChenAAAI 2021 · 5 citations
- Motif-Matching Based Subgraph-Level Attentional Convolutional Network for Graph ClassificationHao Peng, Jianxin Li, Qiran Gong, Yuanxing Ning et al.AAAI 2020 · 75 citations
- DHAKR: Learning Deep Hierarchical Attention-Based Kernelized Representations for Graph ClassificationFeifei Qian, Lu Bai, Lixin Cui, Ming Li et al.AAAI 2025 · 4 citations
- Low-dimensional statistical manifold embedding of directed graphsThorben Funke, Tian Guo, Alen Lancic, Nino Antulov-FantulinICLR 2020 · 6 citations
- Heterogeneous Graph Neural Network via Attribute CompletionDi Jin, Cuiying Huo, Chundong Liang, Liang YangWWW 2021 · 220 citations
