U-BERT: Pre-training User Representations for Improved Recommendation
Zhaopeng Qiu, Xian Wu, Jingyue Gao, Wei Fan
Abstract
Learning user representation is a critical task for recommendation systems as it can encode user preference for personalized services. User representation is generally learned from behavior data, such as clicking interactions and review comments. However, for less popular domains, the behavior data is insufficient to learn precise user representations. To deal with this problem, a natural thought is to leverage content-rich domains to complement user representations. Inspired by the recent success of BERT in NLP, we propose a novel pre-training and fine-tuning based approach U-BERT. Different from typical BERT applications, U-BERT is customized for recommendation and utilizes different frameworks in pre-training and fine-tuning. In pre-training, U-BERT focuses on content-rich domains and introduces a user encoder and a review encoder to model users' behaviors. Two pre-training strategies are proposed to learn the general user representations; In fine-tuning, U-BERT focuses on the target content-insufficient domains. In addition to the user and review encoders inherited from the pre-training stage, U-BERT further introduces an item encoder to model item representations. Besides, a review co-matching layer is proposed to capture more semantic interactions between the reviews of the user and item. Finally, U-BERT combines user representations, item representations and review interaction information to improve recommendation performance. Experiments on six benchmark datasets from different domains demonstrate the state-of-the-art performance of U-BERT.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 6b52b24e-a29e-439b-b9ba-7cece9a1771bCited by top-tier papers20
- ReLLa: Retrieval-enhanced Large Language Models for Lifelong Sequential Behavior Comprehension in RecommendationJianghao Lin, Rong Shan, Chenxu Zhu, Kounianhua Du et al.WWW 2024 · 151 citations
- Conditional Generation Net for Medication RecommendationRui Wu, Zhaopeng Qiu, Jiacheng Jiang, Guilin Qi et al.WWW 2022 · 135 citations
- Exploring Large Language Model for Graph Data Understanding in Online Job RecommendationsLikang Wu, Zhaopeng Qiu, Zhi Zheng, Hengshu Zhu et al.AAAI 2024 · 120 citations
- Harnessing Large Language Models for Text-Rich Sequential RecommendationZhi Zheng, Wenshuo Chao, Zhaopeng Qiu, Hengshu Zhu et al.WWW 2024 · 114 citations
- MISSRec: Pre-training and Transferring Multi-modal Interest-aware Sequence Representation for RecommendationJinpeng Wang, Ziyun Zeng, Yunxiao Wang, Yuting Wang et al.ACM MM 2023 · 62 citations
Builds on1
Related papers
- Towards Multi-Interest Pre-training with Sparse Capsule NetworkZuoli Tang, Lin Wang, Lixin Zou, Xiaolu Zhang et al.SIGIR 2023 · 16 citations
- ActionBert: Leveraging User Actions for Semantic Understanding of User InterfacesZecheng He, Srinivas Sunkara, Xiaoxue Zang, Ying Xu et al.AAAI 2021 · 91 citations
- RecBase: Generative Foundation Model Pretraining for Zero-Shot RecommendationSashuai Zhou, Weinan Gan, Qijiong Liu, Ke Lei et al.EMNLP 2025 · 1 citation
- Improving AMR Parsing with Sequence-to-Sequence Pre-trainingDongqin Xu, Junhui Li, Muhua Zhu, Min Zhang et al.EMNLP 2020 · 57 citations
- Towards Universal Sequence Representation Learning for Recommender SystemsYupeng Hou, Shanlei Mu, Wayne Xin Zhao, Yaliang Li et al.KDD 2022 · 245 citations
