Beyond User Embedding Matrix: Learning to Hash for Modeling Large-Scale Users in Recommendation
Shaoyun Shi, Weizhi Ma, Min Zhang, Yongfeng Zhang, Xinxing Yu, Houzhi Shan, Yiqun Liu, Shaoping Ma
Abstract
Modeling large scale and rare-interaction users are the two major challenges in recommender systems, which derives big gaps between researches and applications. Facing to millions or even billions of users, it is hard to store and leverage personalized preferences with a user embedding matrix in real scenarios. And many researches pay attention to users with rich histories, while users with only one or several interactions are the biggest part in real systems. Previous studies make efforts to handle one of the above issues but rarely tackle efficiency and cold-start problems together.
In this work, a novel user preference representation called Preference Hash (PreHash) is proposed to model large scale users, including rare-interaction ones. In PreHash, a series of buckets are generated based on users' historical interactions. Users with similar preferences are assigned into the same buckets automatically, including warm and cold ones. Representations of the buckets are learned accordingly. Contributing to the designed hash buckets, only limited parameters are stored, which saves a lot of memory for more efficient modeling. Furthermore, when new interactions are made by a user, his buckets and representations will be dynamically updated, which enables more effective understanding and modeling of the user. It is worth mentioning that PreHash is flexible to work with various recommendation algorithms by taking the place of previous user embedding matrices. We combine it with multiple state-of-the-art recommendation methods and conduct various experiments. Comparative results on public datasets show that it not only improves the recommendation performance but also significantly reduces the number of model parameters. To summarize, PreHash has achieved significant improvements in both efficiency and effectiveness for recommender systems.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers8
- Learning Vector-Quantized Item Representation for Transferable Sequential RecommendersYupeng Hou, Zhankui He, Julian J. McAuley, Wayne Xin ZhaoWWW 2023 · 256 citations
- Set2setRank: Collaborative Set to Set Ranking for Implicit Feedback based RecommendationLei Chen, Le Wu, Kun Zhang, Richang Hong et al.SIGIR 2021 · 19 citations
- Long-Tail HashingYong Chen, Yuqing Hou, Shu Leng, Qing Zhang et al.SIGIR 2021 · 17 citations
- Large-scale Comb-K RecommendationHouye Ji, Junxiong Zhu, Chuan Shi, Xiao Wang et al.WWW 2021 · 14 citations
- Clustered Embedding Learning for Recommender SystemsYizhou Chen, Guangda Huzhang, Anxiang Zeng, Qingtao Yu et al.WWW 2023 · 13 citations
Related papers
- Hybrid Embedding Framework for Memory-Efficient Recommendation SystemsSeung Jin Yang, Hyuk-Jae Lee, Chae-Eun RheeDAC 2025
- Multi-Feature Discrete Collaborative Filtering for Fast Cold-Start RecommendationYang Xu, Lei Zhu, Zhiyong Cheng, Jingjing Li et al.AAAI 2020 · 29 citations
- M2EU: Meta Learning for Cold-start Recommendation via Enhancing User Preference EstimationZhenchao Wu, Xiao ZhouSIGIR 2023 · 22 citations
- A Dynamic Meta-Learning Model for Time-Sensitive Cold-Start RecommendationsKrishna Prasad Neupane, Ervine Zheng, Yu Kong, Qi YuAAAI 2022 · 16 citations
- PULSE: Socially-Aware User Representation Modeling Toward Parameter-Efficient Graph Collaborative FilteringDoyun Choi, Cheonwoo Lee, Biniyam Aschalew Tolera, Taewook Ham et al.WWW 2026
