Cluster-wise Graph Transformer with Dual-granularity Kernelized Attention
Siyuan Huang, Yunchong Song, Jiayue Zhou, Zhouhan Lin
摘要
In the realm of graph learning, there is a category of methods that conceptualize graphs as hierarchical structures, utilizing node clustering to capture broader structural information. While generally effective, these methods often rely on a fixed graph coarsening routine, leading to overly homogeneous cluster representations and loss of node-level information. In this paper, we envision the graph as a network of interconnected node sets without compressing each cluster into a single embedding. To enable effective information transfer among these node sets, we propose the Node-to-Cluster Attention (N2C-Attn) mechanism. N2C-Attn incorporates techniques from Multiple Kernel Learning into the kernelized attention framework, effectively capturing information at both node and cluster levels. We then devise an efficient form for N2C-Attn using the cluster-wise message-passing framework, achieving linear time complexity. We further analyze how N2C-Attn combines bi-level feature maps of queries and keys, demonstrating its capability to merge dual-granularity information. The resulting architecture, Cluster-wise Graph Transformer (Cluster-GT), which uses node clusters as tokens and employs our proposed N2C-Attn module, shows superior performance on various graph-level tasks. Code is available at https://github.com/LUMIA-Group/Cluster-wise-Graph-Transformer.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- Unifying and Enhancing Graph Transformers via a Hierarchical Mask FrameworkYujie Xing, Xiao Wang, Bin Wu, Hai Huang 等NeurIPS 2025 · 被引用 2 次
- Can Classic GNNs Be Strong Baselines for Graph-level Tasks? Simple Architectures Meet ExcellenceYuankai Luo, Lei Shi, Xiao-Ming WuICML 2025
- Dual-Kernel Graph Community Contrastive LearningXiang Chen, Kun Yue, Wenjie Liu, Zhenyu Zhang 等AAAI 2026
- From atom to space: A region-based readout function for spatial properties of materialsJiawen Zou, Weimin Tan, Zhongyao Wang, Hao Qi 等ICLR 2026
它引用的顶会 Paper22
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- Open Graph Benchmark: Datasets for Machine Learning on GraphsWeihua Hu, Matthias Fey, Marinka Zitnik, Yuxiao Dong 等NeurIPS 2020 · 被引用 3,935 次
- Transformers are RNNs: Fast Autoregressive Transformers with Linear AttentionAngelos Katharopoulos, Apoorv Vyas, Nikolaos Pappas, François FleuretICML 2020 · 被引用 2,665 次
- Do Transformers Really Perform Badly for Graph Representation?Chengxuan Ying, Tianle Cai, Shengjie Luo, Shuxin Zheng 等NeurIPS 2021 · 被引用 1,632 次
- Recipe for a General, Powerful, Scalable Graph TransformerLadislav Rampásek, Michael Galkin, Vijay Prakash Dwivedi, Anh Tuan Luu 等NeurIPS 2022 · 被引用 1,216 次
相关 Paper
- Deep Multi-modal Graph Clustering via Graph Transformer NetworkQianqian Wang, Haiming Xu, Zihao Zhang, Wei Feng 等AAAI 2025 · 被引用 4 次
- Anchor-Driven Nyström for Deep Graph-Level ClusteringJiaxin Wang, Wenxuan Tu, Lingren Wang, Jieren Cheng 等AAAI 2026
- A Unified Graph Clustering NetworkRenda Han, Xiaobao Wang, Longbiao Wang, Wenxin Zhang 等WWW 2026
- ClusterGNN: Cluster-based Coarse-to-Fine Graph Neural Network for Efficient Feature MatchingYan Shi, Junxiong Cai, Yoli Shavit, Tai-Jiang Mu 等CVPR 2022 · 被引用 91 次
- Attention-driven Graph Clustering NetworkZhihao Peng, Hui Liu, Yuheng Jia, Junhui HouACM MM 2021 · 被引用 135 次
