Memory Disagreement: A Pseudo-Labeling Measure from Training Dynamics for Semi-supervised Graph Learning
Hongbin Pei, Yuheng Xiong, Pinghui Wang, Jing Tao, Jialun Liu, Huiqi Deng, Jie Ma, Xiaohong Guan
摘要
In the realm of semi-supervised graph learning, pseudo-labeling is a pivotal strategy to utilize both labeled and unlabeled nodes for model training. Currently, confidence score is the most frequently used pseudo-labeling measure, however, it suffers from poor calibration and issues in out-of-distribution data. In this paper, we propose memory disagreement (MoDis for short), a novel uncertainty measure for pseudo-labeling. We uncover that training dynamics offer significant insights into prediction uncertainty --- if a graph model makes consistent predictions for an unlabeled node throughout training, the corresponding predicted label is likely to be correct. Thus, the node should be suitable for pseudo-labeling. The basic idea is supported by recent studies on training dynamics. We implement MoDis as the entropy of an accumulated distribution that summarizes the disagreement of the model's predictions throughout training. We further enhance and analyze MoDis in case studies, which show nodes with low MoDis are suitable for pseudo-labeling as these nodes tend to be distant from boundaries in both graph and representation space. We design MoDis based pseudo-label selection algorithm and corresponding pseudo-labeling algorithm, which are applicable to various graph neural networks. We empirically validate MoDis on eight benchmark graph datasets. The experimental results show that pseudo labels given by MoDis have better quality in correctness and information gain, and the algorithm benefits various graph neural networks, achieving an average relative improvement of 3.11% and reaching up to 30.24% when compared to the wildly-used uncertainty measure, confidence score. Moreover, we demonstrate the efficacy of MoDis on out-of-distribution nodes.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
引用它的顶会 Paper7
- DS-Agent: Automated Data Science by Empowering Large Language Models with Case-Based ReasoningSiyuan Guo, Cheng Deng, Ying Wen, Hechang Chen 等ICML 2024 · 被引用 107 次
- GPFedRec: Graph-Guided Personalization for Federated RecommendationChunxu Zhang, Guodong Long, Tianyi Zhou, Zijian Zhang 等KDD 2024 · 被引用 26 次
- Multi-Track Message Passing: Tackling Oversmoothing and Oversquashing in Graph Learning via Preventing Heterophily MixingHongbin Pei, Yu Li, Huiqi Deng, Jingxin Hai 等ICML 2024 · 被引用 19 次
- Semi-supervised Node Importance Estimation with Informative Distribution Modeling for Uncertainty RegularizationYankai Chen, Taotao Wang, Yixiang Fang, Yunyu XiaoWWW 2025 · 被引用 8 次
- Positive and Unlabeled Learning with Controlled Probability Boundary FenceChangchun Li, Yuanchao Dai, Lei Feng, Ximing Li 等ICML 2024 · 被引用 8 次
相关 Paper
- Deep Insights into Noisy Pseudo Labeling on Graph DataBotao Wang, Jia Li, Yang Liu, Jiashun Cheng 等NeurIPS 2023 · 被引用 24 次
- Non-Stationary Predictions May Be More Informative: Exploring Pseudo-Labels with a Two-Phase Pattern of Training DynamicsHongbin Pei, Jingxin Hai, Yu Li, Huiqi Deng 等ICML 2025
- Divide and Denoise: Empowering Simple Models for Robust Semi-Supervised Node Classification against Label NoiseKaize Ding, Xiaoxiao Ma, Yixin Liu, Shirui PanKDD 2024 · 被引用 8 次
- Be Confident! Towards Trustworthy Graph Neural Networks via Confidence CalibrationXiao Wang, Hongrui Liu, Chuan Shi, Cheng YangNeurIPS 2021 · 被引用 158 次
- Identifying and Correcting Label Noise for Robust GNNs via Influence ContradictionWei Ju, Wei Zhang, Siyu Yi, Zhengyang Mao 等ICML 2026
