Reliable Data Distillation on Graph Convolutional Network
Wentao Zhang, Xupeng Miao, Yingxia Shao, Jiawei Jiang, Lei Chen, Olivier Ruas, Bin Cui
Abstract
Graph Convolutional Network (GCN) is a widely used method for learning from graph-based data. However, it fails to use the unlabeled data to its full potential, thereby hindering its ability. Given some pseudo labels of the unlabeled data, the GCN can benefit from this extra supervision. Based on Knowledge Distillation and Ensemble Learning, lots of methods use a teacher-student architecture to make better use of the unlabeled data and then make a better prediction. However, these methods introduce unnecessary training costs and a high bias of student model if the teacher's predictions are unreliable. Besides, the final ensemble gains are limited due to limited diversity in the combined models. Therefore, we propose Reliable Data Distillation, a reliable data driven semi-supervised GCN training method. By defining the node reliability and edge reliability in a graph, we can make better use of high quality data and improve the graph representation learning. Furthermore, considering the data reliability and data importance, we propose a new ensemble learning method for GCN and a novel Self-Boosting SSL Framework to combine the above optimizations. Finally, our extensive evaluation of Reliable Data Distillation on real-world datasets shows that our approach outperforms the state-of-the-art methods on semi-supervised node classification tasks.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 7023bf71-7fb6-4f56-bae4-ccb350af422bCited by top-tier papers31
- Extract the Knowledge of Graph Neural Networks and Go Beyond it: An Effective Knowledge Distillation FrameworkCheng Yang, Jiawei Liu, Chuan ShiWWW 2021 · 153 citations
- Node Dependent Local Smoothing for Scalable Graph LearningWentao Zhang, Mingyu Yang, Zeang Sheng, Yang Li et al.NeurIPS 2021 · 87 citations
- SANCUS: Staleness-Aware Communication-Avoiding Full-Graph Decentralized Training in Large-Scale Graph Neural NetworksJingshu Peng, Zhao Chen, Yingxia Shao, Yanyan Shen et al.VLDB 2022 · 76 citations
- Linkless Link Prediction via Relational DistillationZhichun Guo, William Shiao, Shichang Zhang, Yozen Liu et al.ICML 2023 · 60 citations
- Knowledge Distillation Improves Graph Structure Augmentation for Graph Neural NetworksLirong Wu, Haitao Lin, Yufei Huang, Stan Z. LiNeurIPS 2022 · 60 citations
Related papers
- Divide and Denoise: Empowering Simple Models for Robust Semi-Supervised Node Classification against Label NoiseKaize Ding, Xiaoxiao Ma, Yixin Liu, Shirui PanKDD 2024 · 8 citations
- Collaborative Graph Convolutional Networks: Unsupervised Learning Meets Semi-Supervised LearningBinyuan Hui, Pengfei Zhu, Qinghua HuAAAI 2020 · 67 citations
- From Coarse to Fine: Enable Comprehensive Graph Self-supervised Learning with Multi-granular Semantic EnsembleQianlong Wen, Mingxuan Ju, Zhongyu Ouyang, Chuxu Zhang et al.ICML 2024 · 9 citations
- Noisy Node Classification by Bi-level Optimization Based Multi-Teacher DistillationYujing Liu, Zongqian Wu, Zhengyu Lu, Ci Nie et al.AAAI 2025 · 3 citations
- DualGraph: Improving Semi-supervised Graph Classification via Dual Contrastive LearningXiao Luo, Wei Ju, Meng Qu, Chong Chen et al.ICDE 2022 · 44 citations
