Leveraging Large Language Models for Effective Label-free Node Classification in Text-Attributed Graphs
Taiyan Zhang, Renchi Yang, Yurui Lai, Mingyu Yan, Xiaochun Ye, Dongrui Fan
摘要
Graph neural networks (GNNs) have become the preferred models for node classification in graph data due to their robust capabilities in integrating graph structures and attributes. However, these models heavily depend on a substantial amount of high-quality labeled data for training, which is often costly to obtain. With the rise of large language models (LLMs), a promising approach is to utilize their exceptional zero-shot capabilities and extensive knowledge for node labeling. Despite encouraging results, this approach either requires numerous queries to LLMs or suffers from reduced performance due to noisy labels generated by LLMs. To address these challenges, we introduce Locle, an active self-training framework that does Label-free nOde Classification with LLMs cost-Effectively. Locle iteratively identifies small sets of ''critical'' samples using GNNs and extracts informative pseudo-labels for them with both LLMs and GNNs, serving as additional supervision signals to enhance model training. Specifically, Locle comprises three key components: (i) an effective active node selection strategy for initial annotations; (ii) a careful sample selection scheme to identify ''critical'' nodes based on label disharmonicity and entropy; and (iii) a label refinement module that combines LLMs and GNNs with a rewired topology. Extensive experiments on five benchmark text-attributed graph datasets demonstrate that Locle significantly outperforms state-of-the-art methods under the same query budget to LLMs in terms of label-free node classification. Notably, on the DBLP dataset with 14.3k nodes, Locle achieves an 8.08% improvement in accuracy over the state-of-the-art at a cost of less than one cent. Our code is available at https://github.com/HKBU-LAGAS/Locle.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- ERAlign: Energy-based Representation Alignment of GNNs and LLMs on Text-attributed GraphsXianlin Zeng, Fan Xia, Xiangyu ChenICML 2026
- Both Topology and Text Matter: Revisiting LLM-guided Out-of-Distribution Detection on Text-attributed GraphsYinlin Zhu, Di Wu, Xu Wang, Guocong Quan 等KDD 2026
- When LLMs Encounter Open-world Graph Learning: A Fresh View on Unlabeled Data UncertaintyYanzhe Wen, Xunkai Li, Qi Zhang, Lei Zhu 等ICML 2026
- Bridging Structure and Semantics: Uncertainty-Modulated Dual-Path Diffusion for Robust Text-Attributed Graph LearningZhizhi Yu, Jiachen Liu, Qingyu Li, Dongxiao He 等ICML 2026
它引用的顶会 Paper22
- Simple and Deep Graph Convolutional NetworksMing Chen, Zhewei Wei, Zengfeng Huang, Bolin Ding 等ICML 2020 · 被引用 1,910 次
- Self-Consistency Improves Chain of Thought Reasoning in Language ModelsXuezhi Wang, Jason Wei, Dale Schuurmans, Quoc V. Le 等ICLR 2023 · 被引用 681 次
- LinkBERT: Pretraining Language Models with Document LinksMichihiro Yasunaga, Jure Leskovec, Percy LiangACL 2022 · 被引用 463 次
- Deep Graph Clustering via Dual Correlation ReductionYue Liu, Wenxuan Tu, Sihang Zhou, Xinwang Liu 等AAAI 2022 · 被引用 300 次
- Deep Fusion Clustering NetworkWenxuan Tu, Sihang Zhou, Xinwang Liu, Xifeng Guo 等AAAI 2021 · 被引用 264 次
相关 Paper
- Label-free Node Classification on Graphs with Large Language Models (LLMs)Zhikai Chen, Haitao Mao, Hongzhi Wen, Haoyu Han 等ICLR 2024 · 被引用 103 次
- LLMs Are Noisy Oracles! LLM-based Noise-aware Graph Active Learning for Node ClassificationZeang Sheng, Weiyang Guo, Yingxia Shao, Wentao Zhang 等KDD 2025 · 被引用 1 次
- Leveraging Large Language Models for Node Generation in Few-Shot Learning on Text-Attributed GraphsJianxiang Yu, Yuxiang Ren, Chenghua Gong, Jiaqi Tan 等AAAI 2025 · 被引用 32 次
- Preference-driven Knowledge Distillation for Few-shot Node ClassificationXing Wei, Chunchun Chen, Rui Fan, Xiaofeng Cao 等NeurIPS 2025 · 被引用 2 次
- Dynamic Bundling with Large Language Models for Zero-Shot Inference on Text-Attributed GraphsYusheng Zhao, Qixin Zhang, Xiao Luo, Weizhi Zhang 等NeurIPS 2025 · 被引用 4 次
