Heterformer: Transformer-based Deep Node Representation Learning on Heterogeneous Text-Rich Networks
Bowen Jin, Yu Zhang, Qi Zhu, Jiawei Han
摘要
Representation learning on networks aims to derive a meaningful vector representation for each node, thereby facilitating downstream tasks such as link prediction, node classification, and node clustering. In heterogeneous text-rich networks, this task is more challenging due to (1) presence or absence of text: Some nodes are associated with rich textual information, while others are not; (2) diversity of types: Nodes and edges of multiple types form a heterogeneous network structure. As pretrained language models (PLMs) have demonstrated their effectiveness in obtaining widely generalizable text representations, a substantial amount of effort has been made to incorporate PLMs into representation learning on text-rich networks. However, few of them can jointly consider heterogeneous structure (network) information as well as rich textual semantic information of each node effectively. In this paper, we propose Heterformer, a Heterogeneous Network-Empowered Transformer that performs contextualized text encoding and heterogeneous structure encoding in a unified model. Specifically, we inject heterogeneous structure information into each Transformer layer when encoding node texts. Meanwhile, Heterformer is capable of characterizing node/edge type heterogeneity and encoding nodes with or without texts. We conduct comprehensive experiments on three tasks (i.e., link prediction, node classification, and node clustering) on three large-scale datasets from different domains, where Heterformer outperforms competitive baselines significantly and consistently. The code can be found at https://github.com/PeterGriffinJin/Heterformer.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper7
- RAGraph: A General Retrieval-Augmented Graph Learning FrameworkXinke Jiang, Rihong Qiu, Yongxin Xu, Wentao Zhang 等NeurIPS 2024 · 被引用 42 次
- Patton: Language Model Pretraining on Text-Rich NetworksBowen Jin, Wentao Zhang, Yu Zhang, Yu Meng 等ACL 2023 · 被引用 14 次
- Unifying Text Semantics and Graph Structures for Temporal Text-attributed Graphs with Large Language ModelsSiwei Zhang, Yun Xiong, Yateng Tang, Jiarong Xu 等NeurIPS 2025 · 被引用 9 次
- Instruction-based Hypergraph PretrainingMingdai Yang, Zhiwei Liu, Liangwei Yang, Xiaolong Liu 等SIGIR 2024 · 被引用 4 次
- Can Graph Neural Networks Learn Language with Extremely Weak Text Supervision?Zihao Li, Lecheng Zheng, Bowen Jin, Dongqi Fu 等ACL 2025
它引用的顶会 Paper14
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah 等NeurIPS 2020 · 被引用 64,255 次
- ELECTRA: Pre-training Text Encoders as Discriminators Rather Than GeneratorsKevin Clark, Minh-Thang Luong, Quoc V. Le, Christopher D. ManningICLR 2020 · 被引用 541 次
- Multi-behavior Recommendation with Graph Convolutional NetworksBowen Jin, Chen Gao, Xiangnan He, Depeng Jin 等SIGIR 2020 · 被引用 420 次
- Are we really making much progress?: Revisiting, benchmarking and refining heterogeneous graph neural networksQingsong Lv, Ming Ding, Qiang Liu, Yuxiang Chen 等KDD 2021 · 被引用 249 次
- Dense Passage Retrieval for Open-Domain Question AnsweringVladimir Karpukhin, Barlas Oguz, Sewon Min, Patrick Lewis 等EMNLP 2020 · 被引用 142 次
相关 Paper
- Edgeformers: Graph-Empowered Transformers for Representation Learning on Textual-Edge NetworksBowen Jin, Yu Zhang, Yu Meng, Jiawei HanICLR 2023 · 被引用 5 次
- Bootstrapping Heterogeneous Graph Representation Learning via Large Language Models: A Generalized ApproachHang Gao, Chenhao Zhang, Fengge Wu, Changwen Zheng 等AAAI 2025 · 被引用 6 次
- GraphFormers: GNN-nested Transformers for Representation Learning on Textual GraphJunhan Yang, Zheng Liu, Shitao Xiao, Chaozhuo Li 等NeurIPS 2021 · 被引用 262 次
- Leveraging Contrastive Learning for Enhanced Node Representations in Tokenized Graph TransformersJinsong Chen, Hanpeng Liu, John E. Hopcroft, Kun HeNeurIPS 2024 · 被引用 23 次
- MetaFill: Text Infilling for Meta-Path Generation on Heterogeneous Information NetworksZequn Liu, Kefei Duan, Junwei Yang, Hanwen Xu 等EMNLP 2022 · 被引用 1 次
