Beyond Single Labels: Improving Conversational Recommendation through LLM-Powered Data Augmentation
Haozhe Xu, Xiaohua Wang, Changze Lv, Xiaoqing Zheng
Abstract
Conversational recommender systems (CRSs) enhance recommendation quality by engaging users in multi-turn dialogues, capturing nuanced preferences through natural language interactions. However, these systems often face the false negative issue, where items that a user might like are incorrectly labeled as negative during training, leading to suboptimal recommendations. Expanding the label set through data augmentation presents an intuitive solution but faces the challenge of balancing two key aspects: ensuring semantic relevance and preserving the collaborative information inherent in CRS datasets. To address these issues, we propose a novel data augmentation framework that first leverages an LLM-based semantic retriever to identify diverse and semantically relevant items, which are then filtered by a relevance scorer to remove noisy candidates. Building on this, we introduce a two-stage training strategy balancing semantic relevance and collaborative information. Extensive experiments on two benchmark datasets and user simulators demonstrate significant and consistent performance improvements across various recommenders, highlighting the effectiveness of our approach in advancing CRS performance. 1
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 5cd9d109-3b49-494a-8ff0-51945f696a56Cited by top-tier papers1
Ask how each one uses itBuilds on12
- Chain-of-Thought Prompting Elicits Reasoning in Large Language ModelsJason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma et al.NeurIPS 2022 · 22,562 citations
- LoRA: Low-Rank Adaptation of Large Language ModelsEdward J. Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu et al.ICLR 2022 · 18,833 citations
- BART: Denoising Sequence-to-Sequence Pre-training for Natural Language Generation, Translation, and ComprehensionMike Lewis, Yinhan Liu, Naman Goyal, Marjan Ghazvininejad et al.ACL 2020 · 1,224 citations
- Improving Conversational Recommender Systems via Knowledge Graph based Semantic FusionKun Zhou, Wayne Xin Zhao, Shuqing Bian, Yuanhang Zhou et al.KDD 2020 · 309 citations
- Collaborative Large Language Model for Recommender SystemsYaochen Zhu, Liang Wu, Qi Guo, Liangjie Hong et al.WWW 2024 · 150 citations
Related papers
- Improving Conversational Recommendation Systems via Counterfactual Data SimulationXiaolei Wang, Kun Zhou, Xinyu Tang, Wayne Xin Zhao et al.KDD 2023 · 12 citations
- COLA: Improving Conversational Recommender Systems by Collaborative AugmentationDongding Lin, Jian Wang, Wenjie LiAAAI 2023 · 27 citations
- Broadening the View: Demonstration-augmented Prompt Learning for Conversational RecommendationHuy Dao, Yang Deng, Dung D. Le, Lizi LiaoSIGIR 2024 · 19 citations
- Collaborative Retrieval for Large Language Model-based Conversational Recommender SystemsYaochen Zhu, Chao Wan, Harald Steck, Dawen Liang et al.WWW 2025 · 15 citations
- Generalizing Conversational Dense Retrieval via LLM-Cognition Data AugmentationHaonan Chen, Zhicheng Dou, Kelong Mao, Jiongnan Liu et al.ACL 2024 · 10 citations
