Quantize Sequential Recommenders Without Private Data
Lingfeng Shi, Yuang Liu, Jun Wang, Wei Zhang
Abstract
Deep neural networks have achieved great success in sequential recommendation systems. While maintaining high competence in user modeling and next-item recommendation, these models have long been plagued by the numerous parameters and computation, which inhibit them to be deployed on resource-constrained mobile devices. Model quantization, as one of the main paradigms for compression techniques, converts float parameters to low-bit values to reduce parameter redundancy and accelerate inference. To avoid drastic performance degradation, it usually requests a fine-tuning phase with an original dataset. However, the training set of user-item interactions is not always available due to transmission limits or privacy concerns. In this paper, we propose a novel framework to quantize sequential recommenders without access to any real private data. A generator is employed in the framework to synthesize fake sequence samples to feed the quantized sequential recommendation model and minimize the gap with a full-precision sequential recommendation model. The generator and the quantized model are optimized with a min-max game — alternating discrepancy estimation and knowledge transfer. Moreover, we devise a two-level discrepancy modeling strategy to transfer information between the quantized model and the full-precision model. The extensive experiments of various recommendation networks on three public datasets demonstrate the effectiveness of the proposed framework.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext be1e302d-e617-4049-89e3-4b204c321053Cited by top-tier papers1
Ask how each one uses itBuilds on8
- Next Point-of-Interest Recommendation on Resource-Constrained Mobile DevicesQinyong Wang, Hongzhi Yin, Tong Chen, Zi Huang et al.WWW 2020 · 116 citations
- Learnable Embedding sizes for Recommender SystemsSiyi Liu, Chen Gao, Yihong Chen, Depeng Jin et al.ICLR 2021 · 97 citations
- Student Customized Knowledge Distillation: Bridging the Gap Between Student and TeacherYichen Zhu, Yi WangICCV 2021 · 95 citations
- On-Device Next-Item Recommendation with Self-Supervised Knowledge DistillationXin Xia, Hongzhi Yin, Junliang Yu, Qinyong Wang et al.SIGIR 2022 · 62 citations
- DeepRec: On-device Deep Learning for Privacy-Preserving Sequential Recommendation in Mobile CommerceJialiang Han, Yun Ma, Qiaozhu Mei, Xuanzhe LiuWWW 2021 · 61 citations
Related papers
- Zero-Shot Adversarial QuantizationYuang Liu, Wei Zhang, Jun WangCVPR 2021
- A Generic Network Compression Framework for Sequential Recommender SystemsYang Sun, Fajie Yuan, Min Yang, Guoao Wei et al.SIGIR 2020 · 52 citations
- Causal-DFQ: Causality Guided Data-free Network QuantizationYuzhang Shang, Bingxin Xu, Gaowen Liu, Ramana Rao Kompella et al.ICCV 2023 · 8 citations
- Generative Session-based RecommendationZhidan Wang, Wenwen Ye, Xu Chen, Wenqiang Zhang et al.WWW 2022 · 13 citations
- Sim4Rec: Data-Free Model Extraction Attack on Sequential RecommendationYihao Wang, Jiajie Su, Chaochao Chen, Meng Han et al.AAAI 2025 · 7 citations
