Topology Distillation for Recommender System
SeongKu Kang, Junyoung Hwang, Wonbin Kweon, Hwanjo Yu
Abstract
Recommender Systems (RS) have employed knowledge distillation which is a model compression technique training a compact student model with the knowledge transferred from a pre-trained large teacher model. Recent work has shown that transferring knowledge from the teacher's intermediate layer significantly improves the recommendation quality of the student. However, they transfer the knowledge of individual representation point-wise and thus have a limitation in that primary information of RS lies in the relations in the representation space. This paper proposes a new topology distillation approach that guides the student by transferring the topological structure built upon the relations in the teacher space. We first observe that simply making the student learn the whole topological structure is not always effective and even degrades the student's performance. We demonstrate that because the capacity of the student is highly limited compared to that of the teacher, learning the whole topological structure is daunting for the student. To address this issue, we propose a novel method named Hierarchical Topology Distillation (HTD) which distills the topology hierarchically to cope with the large capacity gap. Our extensive experiments on real-world datasets show that the proposed method significantly outperforms the state-of-the-art competitors. We also provide in-depth analyses to ascertain the benefit of distilling the topology for RS.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 40640a8a-a0c7-4fa7-9ea4-ae6fe8438d61Cited by top-tier papers12
- Linkless Link Prediction via Relational DistillationZhichun Guo, William Shiao, Shichang Zhang, Yozen Liu et al.ICML 2023 · 60 citations
- Distillation from Heterogeneous Models for Top-K RecommendationSeongKu Kang, Wonbin Kweon, Dongha Lee, Jianxun Lian et al.WWW 2023 · 35 citations
- Geometer: Graph Few-Shot Class-Incremental Learning via Prototype RepresentationBin Lu, Xiaoying Gan, Lina Yang, Weinan Zhang et al.KDD 2022 · 18 citations
- Cooperative Retriever and Ranker in Deep RecommendersXu Huang, Defu Lian, Jin Chen, Zheng Liu et al.WWW 2023 · 17 citations
- Consensus Learning from Heterogeneous Objectives for One-Class Collaborative FilteringSeongku Kang, Dongha Lee, Wonbin Kweon, Junyoung Hwang et al.WWW 2022 · 16 citations
Builds on5
- LightGCN: Simplifying and Powering Graph Convolution Network for RecommendationXiangnan He, Kuan Deng, Xiang Wang, Yan Li et al.SIGIR 2020 · 4,448 citations
- Similarity-Preserving Knowledge DistillationFrederick Tung, Greg MoriICCV 2019 · 1,214 citations
- On Sampled Metrics for Item RecommendationWalid Krichene, Steffen RendleKDD 2020 · 459 citations
- Bootstrapping User and Item Representations for One-Class Collaborative FilteringDongha Lee, SeongKu Kang, Hyunjun Ju, Chanyoung Park et al.SIGIR 2021 · 117 citations
- Bidirectional Distillation for Top-K Recommender SystemWonbin Kweon, SeongKu Kang, Hwanjo YuWWW 2021 · 58 citations
Related papers
- Do Topological Characteristics Help in Knowledge Distillation?Jungeun Kim, Junwon You, Dongjin Lee, Ha Young Kim et al.ICML 2024 · 11 citations
- Complementary Relation Contrastive DistillationJinguo Zhu, Shixiang Tang, Dapeng Chen, Shijie Yu et al.CVPR 2021
- Maximizing the Effectiveness of Larger BERT Models for CompressionWen-Shu Fan, Su Lu, Shangyu Xing, Xin-Chun Li et al.ACL 2025
- SHARP-Distill: A 68× Faster Recommender System with Hypergraph Neural Networks and Language ModelsSaman Forouzandeh, Parham Moradi, Mahdi JaliliICML 2025
- Distilling Knowledge From Graph Convolutional NetworksYiding Yang, Jiayan Qiu, Mingli Song, Dacheng Tao et al.CVPR 2020
