StackRec: Efficient Training of Very Deep Sequential Recommender Models by Iterative Stacking
Jiachun Wang, Fajie Yuan, Jian Chen, Qingyao Wu, Min Yang, Yang Sun, Guoxiao Zhang
摘要
Deep learning has brought great progress for the sequential recommendation (SR) tasks. With advanced network architectures, sequential recommender models can be stacked with many hidden layers, e.g., up to 100 layers on real-world recommendation datasets. Training such a deep network is difficult because it can be computationally very expensive and takes much longer time, especially in situations where there are tens of billions of user-item interactions. To deal with such a challenge, we present StackRec, a simple, yet very effective and efficient training framework for deep SR models by iterative layer stacking. Specifically, we first offer an important insight that hidden layers/blocks in a well-trained deep SR model have very similar distributions. Enlightened by this, we propose the stacking operation on the pre-trained layers/blocks to transfer knowledge from a shallower model to a deep model, then we perform iterative stacking so as to yield a much deeper but easier-to-train SR model. We validate the performance of StackRec by instantiating it with four state-of-the-art SR models in three practical scenarios with real-world datasets. Extensive experiments show that StackRec achieves not only comparable performance, but also substantial acceleration in training time, compared to SR models that are trained from scratch. Codes are available at https://github.com/wangjiachun0426/StackRec.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper6
- EfficientTrain: Exploring Generalized Curriculum Learning for Training Visual BackbonesYulin Wang, Yang Yue, Rui Lu, Tianjiao Liu 等ICCV 2023 · 被引用 39 次
- Automated Progressive Learning for Efficient Training of Vision TransformersChanglin Li, Bohan Zhuang, Guangrun Wang, Xiaodan Liang 等CVPR 2022 · 被引用 28 次
- Learning Recommender Systems with Implicit Feedback via Soft Target EnhancementMingyue Cheng, Fajie Yuan, Qi Liu, Shenyang Ge 等SIGIR 2021 · 被引用 22 次
- A User-Adaptive Layer Selection Framework for Very Deep Sequential Recommender ModelsLei Chen, Fajie Yuan, Jiaxi Yang, Xiang Ao 等AAAI 2021 · 被引用 14 次
- Catalog-Native LLM: Speaking Item-ID dialect with Less Entanglement for RecommendationReza Shirkavand, Xiaokai Wei, Chen Wang, Zheng Hui 等ICLR 2026 · 被引用 4 次
它引用的顶会 Paper10
- ALBERT: A Lite BERT for Self-supervised Learning of Language RepresentationsZhenzhong Lan, Mingda Chen, Sebastian Goodman, Kevin Gimpel 等ICLR 2020 · 被引用 7,418 次
- LightGCN: Simplifying and Powering Graph Convolution Network for RecommendationXiangnan He, Kuan Deng, Xiang Wang, Yan Li 等SIGIR 2020 · 被引用 4,448 次
- Parameter-Efficient Transfer from Sequential Behaviors for User Modeling and RecommendationFajie Yuan, Xiangnan He, Alexandros Karatzoglou, Liguang ZhangSIGIR 2020 · 被引用 155 次
- Future Data Helps Training: Modeling Future Contexts for Session-based RecommendationFajie Yuan, Xiangnan He, Haochuan Jiang, Guibing Guo 等WWW 2020 · 被引用 114 次
- How to Retrain Recommender System?: A Sequential Meta-Learning MethodYang Zhang, Fuli Feng, Chenxu Wang, Xiangnan He 等SIGIR 2020 · 被引用 70 次
相关 Paper
- A Generic Network Compression Framework for Sequential Recommender SystemsYang Sun, Fajie Yuan, Min Yang, Guoao Wei 等SIGIR 2020 · 被引用 52 次
- One Sequential Recommendation Model Pretrained from Synthetic Priors Predicts Multiple DatasetsWoosung Kang, Jiwon Jeong, Jonghyeok Shin, Jeongwhan Choi 等KDD 2026
- Breaking the Bottleneck: User-Specific Optimization and Real-Time Inference Integration for Sequential RecommendationWenjia Xie, Hao Wang, Minghao Fang, Ruize Yu 等KDD 2025
- Towards Automatic Discovering of Deep Hybrid Network Architecture for Sequential RecommendationMingyue Cheng, Zhiding Liu, Qi Liu, Shenyang Ge 等WWW 2022 · 被引用 36 次
- SLMRec: Distilling Large Language Models into Small for Sequential RecommendationWujiang Xu, Qitian Wu, Zujie Liang, Jiaojiao Han 等ICLR 2025
