FreeKD: Free-direction Knowledge Distillation for Graph Neural Networks
Kaituo Feng, Changsheng Li, Ye Yuan, Guoren Wang
摘要
Knowledge distillation (KD) has demonstrated its effectiveness to boost the performance of graph neural networks (GNNs), where its goal is to distill knowledge from a deeper teacher GNN into a shallower student GNN. However, it is actually difficult to train a satisfactory teacher GNN due to the well-known over-parametrized and over-smoothing issues, leading to invalid knowledge transfer in practical applications. In this paper, we propose the first Free-direction Knowledge Distillation framework via Reinforcement learning for GNNs, called FreeKD, which is no longer required to provide a deeper well-optimized teacher GNN. The core idea of our work is to collaboratively build two shallower GNNs in an effort to exchange knowledge between them via reinforcement learning in a hierarchical way. As we observe that one typical GNN model often has better and worse performances at different nodes during training, we devise a dynamic and free-direction knowledge transfer strategy that consists of two levels of actions: 1) node-level action determines the directions of knowledge transfer between the corresponding nodes of two networks; and then 2) structure-level action determines which of the local structures generated by the node-level actions to be propagated. In essence, our FreeKD is a general and principled framework which can be naturally compatible with GNNs of different architectures. Extensive experiments on five benchmark datasets demonstrate our FreeKD outperforms two base GNNs in a large margin, and shows its efficacy to various GNNs. More surprisingly, our FreeKD has comparable or even better performance than traditional KD algorithms that distill knowledge from a deeper and stronger teacher GNN.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper13
- Quantifying the Knowledge in GNNs for Reliable Distillation into MLPsLirong Wu, Haitao Lin, Yufei Huang, Stan Z. LiICML 2023 · 被引用 48 次
- Boosting Graph Neural Networks via Adaptive Knowledge DistillationZhichun Guo, Chunhui Zhang, Yujie Fan, Yijun Tian 等AAAI 2023 · 被引用 48 次
- DREAM: Dual Structured Exploration with Mixup for Open-set Graph Domain AdaptionNan Yin, Mengzhu Wang, Zhenghan Chen, Li Shen 等ICLR 2024 · 被引用 28 次
- Efficient Traffic Prediction Through Spatio-Temporal DistillationQianru Zhang, Xinyi Gao, Haixin Wang, Siu Ming Yiu 等AAAI 2025 · 被引用 22 次
- Keypoint-based Progressive Chain-of-Thought Distillation for LLMsKaituo Feng, Changsheng Li, Xiaolu Zhang, Jun Zhou 等ICML 2024 · 被引用 20 次
它引用的顶会 Paper12
- Simple and Deep Graph Convolutional NetworksMing Chen, Zhewei Wei, Zengfeng Huang, Bolin Ding 等ICML 2020 · 被引用 1,910 次
- Geom-GCN: Geometric Graph Convolutional NetworksHongbin Pei, Bingzhe Wei, Kevin Chen-Chuan Chang, Yu Lei 等ICLR 2020 · 被引用 1,445 次
- Improved Knowledge Distillation via Teacher AssistantSeyed-Iman Mirzadeh, Mehrdad Farajtabar, Ang Li, Nir Levine 等AAAI 2020 · 被引用 1,361 次
- Reinforced Multi-Teacher Selection for Knowledge DistillationFei Yuan, Linjun Shou, Jian Pei, Wutao Lin 等AAAI 2021 · 被引用 155 次
- Extract the Knowledge of Graph Neural Networks and Go Beyond it: An Effective Knowledge Distillation FrameworkCheng Yang, Jiawei Liu, Chuan ShiWWW 2021 · 被引用 153 次
相关 Paper
- Multi-Scale Distillation from Multiple Graph Neural NetworksChunhai Zhang, Jie Liu, Kai Dang, Wenzheng ZhangAAAI 2022 · 被引用 17 次
- Compressing Deep Graph Neural Networks via Adversarial Knowledge DistillationHuarui He, Jie Wang, Zhanqiu Zhang, Feng WuKDD 2022 · 被引用 44 次
- Accelerating Molecular Graph Neural Networks via Knowledge DistillationFilip Ekström Kelvinius, Dimitar Georgiev, Artur P. Toshev, Johannes GasteigerNeurIPS 2023 · 被引用 22 次
- Narrow the Input Mismatch in Deep Graph Neural Network DistillationQiqi Zhou, Yanyan Shen, Lei ChenKDD 2023 · 被引用 5 次
- Pareto-Based Heterogeneous Knowledge Distillation for MLPs on GraphsWenrui Zhao, Yijun Tian, Zhichao Xu, Yawei Wang 等AAAI 2026
