Deep Insights into Noisy Pseudo Labeling on Graph Data
Botao Wang, Jia Li, Yang Liu, Jiashun Cheng, Yu Rong, Wenjia Wang, Fugee Tsung
摘要
Pseudo labeling (PL) is a wide-applied strategy to enlarge the labeled dataset by self-annotating the potential samples during the training process. Several works have shown that it can improve the graph learning model performance in general. However, we notice that the incorrect labels can be fatal to the graph training process. Inappropriate PL may result in the performance degrading, especially on graph data where the noise can propagate. Surprisingly, the corresponding error is seldom theoretically analyzed in the literature. In this paper, we aim to give deep insights of PL on graph learning models. We first present the error analysis of PL strategy by showing that the error is bounded by the confidence of PL threshold and consistency of multi-view prediction. Then, we theoretically illustrate the effect of PL on convergence property. Based on the analysis, we propose a cautious pseudo labeling methodology in which we pseudo label the samples with highest confidence and multi-view consistency. Finally, extensive experiments demonstrate that the proposed strategy improves graph learning process and outperforms other PL strategies on link prediction and node classification tasks.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper7
- Weakly Supervised Anomaly Detection via Knowledge-Data AlignmentHaihong Zhao, Chenyi Zi, Yang Liu, Chen Zhang 等WWW 2024 · 被引用 17 次
- Inductive Attributed Community Search: to Learn Communities across GraphsShuheng Fang, Kangfei Zhao, Yu Rong, Zhixun Li 等VLDB 2024 · 被引用 10 次
- Cross-Domain Graph Data Scaling: A Showcase with Diffusion ModelsWenzhuo Tang, Haitao Mao, Danial Dervovic, Ivan Brugere 等NeurIPS 2025 · 被引用 8 次
- Non-Stationary Predictions May Be More Informative: Exploring Pseudo-Labels with a Two-Phase Pattern of Training DynamicsHongbin Pei, Jingxin Hai, Yu Li, Huiqi Deng 等ICML 2025
- Can Pseudo-Label Be More Reliable? A Simple yet Effective Topology-Aware Graph Self-Training MethodGen Liu, Zhongying Zhao, Hui Zhou, Chao Li 等AAAI 2026
它引用的顶会 Paper16
- FixMatch: Simplifying Semi-Supervised Learning with Consistency and ConfidenceKihyuk Sohn, David Berthelot, Nicholas Carlini, Zizhao Zhang 等NeurIPS 2020 · 被引用 5,129 次
- Confidence Regularized Self-TrainingYang Zou, Zhiding Yu, Xiaofeng Liu, B. V. K. Vijaya Kumar 等ICCV 2019 · 被引用 901 次
- Self-paced Contrastive Learning with Hybrid Memory for Domain Adaptive Object Re-IDYixiao Ge, Feng Zhu, Dapeng Chen, Rui Zhao 等NeurIPS 2020 · 被引用 688 次
- Mutual Mean-Teaching: Pseudo Label Refinery for Unsupervised Domain Adaptation on Person Re-identificationYixiao Ge, Dapeng Chen, Hongsheng LiICLR 2020 · 被引用 651 次
- Self-supervised Co-Training for Video Representation LearningTengda Han, Weidi Xie, Andrew ZissermanNeurIPS 2020 · 被引用 405 次
相关 Paper
- Cross-View Graph Consistency Learning for Invariant Graph RepresentationsJie Chen, Hua Mao, Wai Lok Woo, Chuanbin Liu 等AAAI 2025 · 被引用 1 次
- Memory Disagreement: A Pseudo-Labeling Measure from Training Dynamics for Semi-supervised Graph LearningHongbin Pei, Yuheng Xiong, Pinghui Wang, Jing Tao 等WWW 2024 · 被引用 11 次
- NRGNN: Learning a Label Noise Resistant Graph Neural Network on Sparsely and Noisily Labeled GraphsEnyan Dai, Charu Aggarwal, Suhang WangKDD 2021 · 被引用 80 次
- Uncertainty-Aware Pseudo-Labeling and Dual Graph Driven Network for Incomplete Multi-View Multi-Label ClassificationWulin Xie, Xiaohuan Lu, Yadong Liu, Jiang Long 等ACM MM 2024 · 被引用 17 次
- Nested Graph Pseudo-Label Refinement for Noisy Label Domain Adaptation LearningYingxu Wang, Mengzhu Wang, Zhichao Huang, Suyu Liu 等AAAI 2026 · 被引用 7 次
