Leveraging Large Language Models for Effective Label-free Node Classification in Text-Attributed Graphs
Taiyan Zhang, Renchi Yang, Yurui Lai, Mingyu Yan, Xiaochun Ye, Dongrui Fan
Abstract
Graph neural networks (GNNs) have become the preferred models for node classification in graph data due to their robust capabilities in integrating graph structures and attributes. However, these models heavily depend on a substantial amount of high-quality labeled data for training, which is often costly to obtain. With the rise of large language models (LLMs), a promising approach is to utilize their exceptional zero-shot capabilities and extensive knowledge for node labeling. Despite encouraging results, this approach either requires numerous queries to LLMs or suffers from reduced performance due to noisy labels generated by LLMs. To address these challenges, we introduce Locle, an active self-training framework that does Label-free nOde Classification with LLMs cost-Effectively. Locle iteratively identifies small sets of ''critical'' samples using GNNs and extracts informative pseudo-labels for them with both LLMs and GNNs, serving as additional supervision signals to enhance model training. Specifically, Locle comprises three key components: (i) an effective active node selection strategy for initial annotations; (ii) a careful sample selection scheme to identify ''critical'' nodes based on label disharmonicity and entropy; and (iii) a label refinement module that combines LLMs and GNNs with a rewired topology. Extensive experiments on five benchmark text-attributed graph datasets demonstrate that Locle significantly outperforms state-of-the-art methods under the same query budget to LLMs in terms of label-free node classification. Notably, on the DBLP dataset with 14.3k nodes, Locle achieves an 8.08% improvement in accuracy over the state-of-the-art at a cost of less than one cent. Our code is available at https://github.com/HKBU-LAGAS/Locle.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 53e63eb6-69cd-415d-8ff8-131835eada60Cited by top-tier papers4
- ERAlign: Energy-based Representation Alignment of GNNs and LLMs on Text-attributed GraphsXianlin Zeng, Fan Xia, Xiangyu ChenICML 2026
- Both Topology and Text Matter: Revisiting LLM-guided Out-of-Distribution Detection on Text-attributed GraphsYinlin Zhu, Di Wu, Xu Wang, Guocong Quan et al.KDD 2026
- When LLMs Encounter Open-world Graph Learning: A Fresh View on Unlabeled Data UncertaintyYanzhe Wen, Xunkai Li, Qi Zhang, Lei Zhu et al.ICML 2026
- Bridging Structure and Semantics: Uncertainty-Modulated Dual-Path Diffusion for Robust Text-Attributed Graph LearningZhizhi Yu, Jiachen Liu, Qingyu Li, Dongxiao He et al.ICML 2026
Builds on22
- Simple and Deep Graph Convolutional NetworksMing Chen, Zhewei Wei, Zengfeng Huang, Bolin Ding et al.ICML 2020 · 1,910 citations
- Self-Consistency Improves Chain of Thought Reasoning in Language ModelsXuezhi Wang, Jason Wei, Dale Schuurmans, Quoc V. Le et al.ICLR 2023 · 681 citations
- LinkBERT: Pretraining Language Models with Document LinksMichihiro Yasunaga, Jure Leskovec, Percy LiangACL 2022 · 463 citations
- Deep Graph Clustering via Dual Correlation ReductionYue Liu, Wenxuan Tu, Sihang Zhou, Xinwang Liu et al.AAAI 2022 · 300 citations
- Deep Fusion Clustering NetworkWenxuan Tu, Sihang Zhou, Xinwang Liu, Xifeng Guo et al.AAAI 2021 · 264 citations
Related papers
- Label-free Node Classification on Graphs with Large Language Models (LLMs)Zhikai Chen, Haitao Mao, Hongzhi Wen, Haoyu Han et al.ICLR 2024 · 103 citations
- LLMs Are Noisy Oracles! LLM-based Noise-aware Graph Active Learning for Node ClassificationZeang Sheng, Weiyang Guo, Yingxia Shao, Wentao Zhang et al.KDD 2025 · 1 citation
- Leveraging Large Language Models for Node Generation in Few-Shot Learning on Text-Attributed GraphsJianxiang Yu, Yuxiang Ren, Chenghua Gong, Jiaqi Tan et al.AAAI 2025 · 32 citations
- Preference-driven Knowledge Distillation for Few-shot Node ClassificationXing Wei, Chunchun Chen, Rui Fan, Xiaofeng Cao et al.NeurIPS 2025 · 2 citations
- Dynamic Bundling with Large Language Models for Zero-Shot Inference on Text-Attributed GraphsYusheng Zhao, Qixin Zhang, Xiao Luo, Weizhi Zhang et al.NeurIPS 2025 · 4 citations
