Compressing Deep Graph Neural Networks via Adversarial Knowledge Distillation
Huarui He, Jie Wang, Zhanqiu Zhang, Feng Wu
摘要
Deep graph neural networks (GNNs) have been shown to be expressive for modeling graph-structured data. Nevertheless, the overstacked architecture of deep graph models makes it difficult to deploy and rapidly test on mobile or embedded systems. To compress over-stacked GNNs, knowledge distillation via a teacher-student architecture turns out to be an effective technique, where the key step is to measure the discrepancy between teacher and student networks with predefined distance functions. However, using the same distance for graphs of various structures may be unfit, and the optimal distance formulation is hard to determine. To tackle these problems, we propose a novel Adversarial Knowledge Distillation framework for graph models named GraphAKD, which adversarially trains a discriminator and a generator to adaptively detect and decrease the discrepancy. Specifically, noticing that the well-captured inter-node and inter-class correlations favor the success of deep GNNs, we propose to criticize the inherited knowledge from node-level and class-level views with a trainable discriminator. The discriminator distinguishes between teacher knowledge and what the student inherits, while the student GNN works as a generator and aims to fool the discriminator. To our best knowledge, GraphAKD is the first to introduce adversarial training to knowledge distillation in graph domains. Experiments on nodelevel and graph-level classification benchmarks demonstrate that GraphAKD improves the student performance by a large margin. The results imply that GraphAKD can precisely transfer knowledge from a complicated teacher GNN to a compact student GNN.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper6
- Efficient Traffic Prediction Through Spatio-Temporal DistillationQianru Zhang, Xinyi Gao, Haixin Wang, Siu Ming Yiu 等AAAI 2025 · 被引用 22 次
- Accelerating Molecular Graph Neural Networks via Knowledge DistillationFilip Ekström Kelvinius, Dimitar Georgiev, Artur P. Toshev, Johannes GasteigerNeurIPS 2023 · 被引用 22 次
- Accelerating Scalable Graph Neural Network Inference with Node-Adaptive PropagationXinyi Gao, Wentao Zhang, Junliang Yu, Yingxia Shao 等ICDE 2024 · 被引用 15 次
- Self-Training Based Few-Shot Node Classification by Knowledge DistillationZongqian Wu, Yujie Mo, Peng Zhou, Shangbo Yuan 等AAAI 2024 · 被引用 11 次
- ProtGO: Function-Guided Protein Modeling for Unified Representation LearningBozhen Hu, Cheng Tan, Yongjie Xu, Zhangyang Gao 等NeurIPS 2024 · 被引用 10 次
它引用的顶会 Paper13
- LightGCN: Simplifying and Powering Graph Convolution Network for RecommendationXiangnan He, Kuan Deng, Xiang Wang, Yan Li 等SIGIR 2020 · 被引用 4,448 次
- Open Graph Benchmark: Datasets for Machine Learning on GraphsWeihua Hu, Matthias Fey, Marinka Zitnik, Yuxiao Dong 等NeurIPS 2020 · 被引用 3,935 次
- Simple and Deep Graph Convolutional NetworksMing Chen, Zhewei Wei, Zengfeng Huang, Bolin Ding 等ICML 2020 · 被引用 1,910 次
- GraphSAINT: Graph Sampling Based Inductive Learning MethodHanqing Zeng, Hongkuan Zhou, Ajitesh Srivastava, Rajgopal Kannan 等ICLR 2020 · 被引用 1,155 次
- Adversarial Graph Augmentation to Improve Graph Contrastive LearningSusheel Suresh, Pan Li, Cong Hao, Jennifer NevilleNeurIPS 2021 · 被引用 475 次
相关 Paper
- Boosting Graph Neural Networks via Adaptive Knowledge DistillationZhichun Guo, Chunhui Zhang, Yujie Fan, Yijun Tian 等AAAI 2023 · 被引用 48 次
- Narrow the Input Mismatch in Deep Graph Neural Network DistillationQiqi Zhou, Yanyan Shen, Lei ChenKDD 2023 · 被引用 5 次
- Multi-Scale Distillation from Multiple Graph Neural NetworksChunhai Zhang, Jie Liu, Kai Dang, Wenzheng ZhangAAAI 2022 · 被引用 17 次
- Distilling Portable Generative Adversarial Networks for Image TranslationHanting Chen, Yunhe Wang, Han Shu, Changyuan Wen 等AAAI 2020 · 被引用 89 次
- FreeKD: Free-direction Knowledge Distillation for Graph Neural NetworksKaituo Feng, Changsheng Li, Ye Yuan, Guoren WangKDD 2022 · 被引用 28 次
