A Generic Network Compression Framework for Sequential Recommender Systems
Yang Sun, Fajie Yuan, Min Yang, Guoao Wei, Zhou Zhao, Duo Liu
摘要
Sequential recommender systems (SRS) have become the key technology in capturing user's dynamic interests and generating highquality recommendations. Current state-of-the-art sequential recommender models are typically based on a sandwich-structured deep neural network, where one or more middle (hidden) layers are placed between the input embedding layer and output so max layer. In general, these models require a large number of parameters to obtain optimal performance. Despite the effectiveness, at some point, further increasing model size may be harder for model deployment in resource-constraint devices. To resolve the issues, we propose a compressed sequential recommendation framework, termed as CpRec, where two generic model shrinking techniques are employed. Specifically, we first propose a block-wise adaptive decomposition to approximate the input and so max matrices by exploiting the fact that items in SRS obey a long-tailed distribution. To reduce the parameters of the middle layers, we introduce three layer-wise parameter sharing schemes. We instantiate CpRec using deep convolutional neural network with dilated kernels given consideration to both recommendation accuracy and efficiency. By the extensive ablation studies, we demonstrate that the proposed CpRec can achieve up to 4∼8 times compression rates in real-world SRS datasets. Meanwhile, CpRec is faster during training & inference, and in most cases outperforms its uncompressed counterpart. Our code is available at h ps://github.com/siat-nlp/CpRec.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper11
- Parameter-Efficient Transfer from Sequential Behaviors for User Modeling and RecommendationFajie Yuan, Xiangnan He, Alexandros Karatzoglou, Liguang ZhangSIGIR 2020 · 被引用 155 次
- Accelerating Recommendation System Training by Leveraging Popular ChoicesMuhammad Adnan, Yassaman Ebrahimzadeh Maboud, Divya Mahajan, Prashant J. NairVLDB 2022 · 被引用 70 次
- On-Device Next-Item Recommendation with Self-Supervised Knowledge DistillationXin Xia, Hongzhi Yin, Junliang Yu, Qinyong Wang 等SIGIR 2022 · 被引用 62 次
- One Person, One Model, One World: Learning Continual User Representation without ForgettingFajie Yuan, Guoxiao Zhang, Alexandros Karatzoglou, Joemon M. Jose 等SIGIR 2021 · 被引用 52 次
- Towards Automatic Discovering of Deep Hybrid Network Architecture for Sequential RecommendationMingyue Cheng, Zhiding Liu, Qi Liu, Shenyang Ge 等WWW 2022 · 被引用 36 次
它引用的顶会 Paper2
- ALBERT: A Lite BERT for Self-supervised Learning of Language RepresentationsZhenzhong Lan, Mingda Chen, Sebastian Goodman, Kevin Gimpel 等ICLR 2020 · 被引用 7,418 次
- LightGCN: Simplifying and Powering Graph Convolution Network for RecommendationXiangnan He, Kuan Deng, Xiang Wang, Yan Li 等SIGIR 2020 · 被引用 4,448 次
相关 Paper
- StackRec: Efficient Training of Very Deep Sequential Recommender Models by Iterative StackingJiachun Wang, Fajie Yuan, Jian Chen, Qingyao Wu 等SIGIR 2021 · 被引用 25 次
- Quantize Sequential Recommenders Without Private DataLingfeng Shi, Yuang Liu, Jun Wang, Wei ZhangWWW 2023 · 被引用 3 次
- A User-Adaptive Layer Selection Framework for Very Deep Sequential Recommender ModelsLei Chen, Fajie Yuan, Jiaxi Yang, Xiang Ao 等AAAI 2021 · 被引用 14 次
- Breaking the Bottleneck: User-Specific Optimization and Real-Time Inference Integration for Sequential RecommendationWenjia Xie, Hao Wang, Minghao Fang, Ruize Yu 等KDD 2025
- SAGE: Global Semantic Alignment with LLMs for Long-Tail Sequential RecommendationMaolin Wang, Tongshu Bian, Ziyan Wang, Xiaotong Jiang 等WWW 2026
