Unified Conversational Recommendation Policy Learning via Graph-based Reinforcement Learning
Yang Deng, Yaliang Li, Fei Sun, Bolin Ding, Wai Lam
摘要
Conversational recommender systems (CRS) enable the traditional recommender systems to explicitly acquire user preferences towards items and attributes through interactive conversations. Reinforcement learning (RL) is widely adopted to learn conversational recommendation policies to decide what attributes to ask, which items to recommend, and when to ask or recommend, at each conversation turn. However, existing methods mainly target at solving one or two of these three decision-making problems in CRS with separated conversation and recommendation components, which restrict the scalability and generality of CRS and fall short of preserving a stable training procedure. In the light of these challenges, we propose to formulate these three decision-making problems in CRS as a unified policy learning task. In order to systematically integrate conversation and recommendation components, we develop a dynamic weighted graph based RL method to learn a policy to select the action at each conversation turn, either asking an attribute or recommending items. Further, to deal with the sample efficiency issue, we propose two action selection strategies for reducing the candidate action space according to the preference and entropy information. Experimental results on two benchmark CRS datasets and a real-world E-Commerce application show that the proposed method not only significantly outperforms state-of-the-art methods but also enhances the scalability and stability of CRS.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper29
- Plug-and-Play Policy Planner for Large Language Model Powered Dialogue AgentsYang Deng, Wenxuan Zhang, Wai Lam, See-Kiong Ng 等ICLR 2024 · 被引用 86 次
- Multiple Choice Questions based Multi-Interest Policy Learning for Conversational RecommendationYiming Zhang, Lingfei Wu, Qi Shen, Yitong Pang 等WWW 2022 · 被引用 72 次
- Learning Neural Templates for Recommender Dialogue SystemZujie Liang, Huang Hu, Can Xu, Jian Miao 等EMNLP 2021 · 被引用 40 次
- User Satisfaction Estimation with Sequential Dialogue Act Modeling in Goal-oriented Conversational SystemsYang Deng, Wenxuan Zhang, Wai Lam, Hong Cheng 等WWW 2022 · 被引用 34 次
- Variational Reasoning about User Preferences for Conversational RecommendationZhaochun Ren, Zhi Tian, Dongdong Li, Pengjie Ren 等SIGIR 2022 · 被引用 31 次
它引用的顶会 Paper11
- LightGCN: Simplifying and Powering Graph Convolution Network for RecommendationXiangnan He, Kuan Deng, Xiang Wang, Yan Li 等SIGIR 2020 · 被引用 4,448 次
- Improving Conversational Recommender Systems via Knowledge Graph based Semantic FusionKun Zhou, Wayne Xin Zhao, Shuqing Bian, Yuanhang Zhou 等KDD 2020 · 被引用 309 次
- Self-Supervised Reinforcement Learning for Recommender SystemsXin Xin, Alexandros Karatzoglou, Ioannis Arapakis, Joemon M. JoseSIGIR 2020 · 被引用 217 次
- Reinforced Negative Sampling over Knowledge Graph for RecommendationXiang Wang, Yaokun Xu, Xiangnan He, Yixin Cao 等WWW 2020 · 被引用 209 次
- Interactive Recommender System via Knowledge Graph-enhanced Reinforcement LearningSijin Zhou, Xinyi Dai, Haokun Chen, Weinan Zhang 等SIGIR 2020 · 被引用 166 次
相关 Paper
- Confident Action Decision via Hierarchical Policy Learning for Conversational RecommendationHeeseon Kim, Hyeongjun Yang, Kyong-Ho LeeWWW 2023 · 被引用 10 次
- Interactive Path Reasoning on Graph for Conversational RecommendationWenqiang Lei, Gangyi Zhang, Xiangnan He, Yisong Miao 等KDD 2020 · 被引用 158 次
- Learning to Infer User Implicit Preference in Conversational RecommendationChenhao Hu, Shuhua Huang, Yansen Zhang, Yubao LiuSIGIR 2022 · 被引用 38 次
- Multi-Objective Intrinsic Reward Learning for Conversational Recommender SystemsZhendong Chu, Nan Wang, Hongning WangNeurIPS 2023 · 被引用 5 次
- HutCRS: Hierarchical User-Interest Tracking for Conversational Recommender SystemMingjie Qian, Yongsen Zheng, Jinghui Qin, Liang LinEMNLP 2023 · 被引用 11 次
