MulDE: Multi-teacher Knowledge Distillation for Low-dimensional Knowledge Graph Embeddings
Kai Wang, Yu Liu, Qian Ma, Quan Z. Sheng
摘要
Link prediction based on knowledge graph embeddings (KGE) aims to predict new triples to automatically construct knowledge graphs (KGs). However, recent KGE models achieve performance improvements by excessively increasing the embedding dimensions, which may cause enormous training costs and require more storage space. In this paper, instead of training high-dimensional models, we propose MulDE, a novel knowledge distillation framework, which includes multiple low-dimensional hyperbolic KGE models as teachers and two student components, namely Junior and Senior. Under a novel iterative distillation strategy, the Junior component, a low-dimensional KGE model, asks teachers actively based on its preliminary prediction results, and the Senior component integrates teachers' knowledge adaptively to train the Junior component based on two mechanisms: relation-specific scaling and contrast attention. The experimental results show that MulDE can effectively improve the performance and training speed of lowdimensional KGE models. The distilled 32-dimensional model is competitive compared to the state-of-the-art high-dimensional methods on several widely-used datasets. CCS CONCEPTS • Computing methodologies → Knowledge representation and reasoning; • Information systems → Entity relationship models.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper19
- NodePiece: Compositional and Parameter-Efficient Representations of Large Knowledge GraphsMikhail Galkin, Etienne G. Denis, Jiapeng Wu, William L. HamiltonICLR 2022 · 被引用 114 次
- Towards Continual Knowledge Graph Embedding via Incremental DistillationJiajun Liu, Wenjun Ke, Peng Wang, Ziyu Shang 等AAAI 2024 · 被引用 52 次
- Boosting Graph Neural Networks via Adaptive Knowledge DistillationZhichun Guo, Chunhui Zhang, Yujie Fan, Yijun Tian 等AAAI 2023 · 被引用 48 次
- Link Prediction with Attention Applied on Multiple Knowledge Graph Embedding ModelsCosimo Gregucci, Mojtaba Nayyeri, Daniel Hernández, Steffen StaabWWW 2023 · 被引用 37 次
- Distillation from Heterogeneous Models for Top-K RecommendationSeongKu Kang, Wonbin Kweon, Dongha Lee, Jianxun Lian 等WWW 2023 · 被引用 35 次
它引用的顶会 Paper14
- Improved Knowledge Distillation via Teacher AssistantSeyed-Iman Mirzadeh, Mehrdad Farajtabar, Ang Li, Nir Levine 等AAAI 2020 · 被引用 1,361 次
- Composition-based Multi-Relational Graph Convolutional NetworksShikhar Vashishth, Soumya Sanyal, Vikram Nitin, Partha P. TalukdarICLR 2020 · 被引用 1,105 次
- You CAN Teach an Old Dog New Tricks! On Training Knowledge Graph EmbeddingsDaniel Ruffinelli, Samuel Broscheit, Rainer GemullaICLR 2020 · 被引用 238 次
- Understanding Knowledge Distillation in Non-autoregressive Machine TranslationChunting Zhou, Jiatao Gu, Graham NeubigICLR 2020 · 被引用 235 次
- Reinforced Negative Sampling over Knowledge Graph for RecommendationXiang Wang, Yaokun Xu, Xiangnan He, Yixin Cao 等WWW 2020 · 被引用 209 次
相关 Paper
- IterDE: An Iterative Knowledge Distillation Framework for Knowledge Graph EmbeddingsJiajun Liu, Peng Wang, Ziyu Shang, Chenxiao WuAAAI 2023 · 被引用 26 次
- Swift and Sure: Hardness-aware Contrastive Learning for Low-dimensional Knowledge Graph EmbeddingsKai Wang, Yu Liu, Quan Z. ShengWWW 2022 · 被引用 21 次
- Geometry Awakening: Cross-Geometry Learning Exhibits Superiority over Individual StructuresYadong Sun, Xiaofeng Cao, Yu Wang, Wei Ye 等NeurIPS 2024 · 被引用 2 次
- Joint Pre-training and Local Re-training: Transferable Representation Learning on Multi-source Knowledge GraphsZequn Sun, Jiacheng Huang, Jinghao Lin, Xiaozhou Xu 等KDD 2023 · 被引用 5 次
- MARCH: Multi-Teacher Contrastive Hypergraph DistillationRongwei Xu, Zitai Qiu, Pengfei Ding, Jia Wu 等WWW 2026
