Neural Interactive Collaborative Filtering
Lixin Zou, Long Xia, Yulong Gu, Xiangyu Zhao, Weidong Liu, Jimmy Xiangji Huang, Dawei Yin
摘要
In this paper, we study collaborative filtering in an interactive setting, in which the recommender agents iterate between making recommendations and updating the user profile based on the interactive feedback. The most challenging problem in this scenario is how to suggest items when the user profile has not been well established, i.e., recommend for cold-start users or warm-start users with taste drifting. Existing approaches either rely on overly pessimistic linear exploration strategy or adopt meta-learning based algorithms in a full exploitation way. In this work, to quickly catch up with the user's interests, we propose to represent the exploration policy with a neural network and directly learn it from the feedback data. Specifically, the exploration policy is encoded in the weights of multi-channel stacked self-attention neural networks and trained with efficient Q-learning by maximizing users' overall satisfaction in the recommender systems. The key insight is that the satisfied recommendations triggered by the exploration recommendation can be viewed as the exploration bonus (delayed reward) for its contribution on improving the quality of the user profile. Therefore, the proposed exploration policy, to balance between learning the user profile and making accurate recommendations, can be directly optimized by maximizing users' long-term satisfaction with reinforcement learning. Extensive experiments and analysis conducted on three benchmark collaborative filtering datasets have demonstrated the advantage of our method over state-of-the-art methods.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper18
- Hypergraph Contrastive Collaborative FilteringLianghao Xia, Chao Huang, Yong Xu, Jiashu Zhao 等SIGIR 2022 · 被引用 445 次
- DEAR: Deep Reinforcement Learning for Online Advertising Impression in Recommender SystemsXiangyu Zhao, Changsheng Gu, Haoshenglun Zhang, Xiwang Yang 等AAAI 2021 · 被引用 131 次
- LinRec: Linear Attention Mechanism for Long-term Sequential Recommender SystemsLangming Liu, Liu Cai, Chi Zhang, Xiangyu Zhao 等SIGIR 2023 · 被引用 86 次
- Multiple Choice Questions based Multi-Interest Policy Learning for Conversational RecommendationYiming Zhang, Lingfei Wu, Qi Shen, Yitong Pang 等WWW 2022 · 被引用 72 次
- AutoDim: Field-aware Embedding Dimension Searchin Recommender SystemsXiangyu Zhao, Haochen Liu, Hui Liu, Jiliang Tang 等WWW 2021 · 被引用 69 次
相关 Paper
- M2EU: Meta Learning for Cold-start Recommendation via Enhancing User Preference EstimationZhenchao Wu, Xiao ZhouSIGIR 2023 · 被引用 22 次
- Context Uncertainty in Contextual Bandits with Applications to Recommender SystemsHao Wang, Yifei Ma, Hao Ding, Yuyang WangAAAI 2022 · 被引用 6 次
- Meta-Learning for Online Update of Recommender SystemsMinseok Kim, Hwanjun Song, Yooju Shin, Dongmin Park 等AAAI 2022 · 被引用 25 次
- PNMTA: A Pretrained Network Modulation and Task Adaptation Approach for User Cold-Start RecommendationHaoyu Pang, Fausto Giunchiglia, Ximing Li, Renchu Guan 等WWW 2022 · 被引用 25 次
- Deployable and Continuable Meta-learning-Based Recommender System with Fast User-Incremental UpdatesRenchu Guan, Haoyu Pang, Fausto Giunchiglia, Ximing Li 等SIGIR 2022 · 被引用 6 次
