Evaluating Conversational Recommender Systems via User Simulation
Shuo Zhang, Krisztian Balog
摘要
Conversational information access is an emerging research area. Currently, human evaluation is used for end-to-end system evaluation, which is both very time and resource intensive at scale, and thus becomes a bottleneck of progress. As an alternative, we propose automated evaluation by means of simulating users. Our user simulator aims to generate responses that a real human would give by considering both individual preferences and the general flow of interaction with the system. We evaluate our simulation approach on an item recommendation task by comparing three existing conversational recommender systems. We show that preference modeling and task-specific interaction models both contribute to more realistic simulations, and can help achieve high correlation between automatic evaluation measures and manual human assessments.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper14
- Rethinking the Evaluation for Conversational Recommendation in the Era of Large Language ModelsXiaolei Wang, Xinyu Tang, Xin Zhao, Jingyuan Wang 等EMNLP 2023 · 被引用 69 次
- An In-depth Investigation of User Response Simulation for Conversational SearchZhenduo Wang, Zhichao Xu, Vivek Srikumar, Qingyao AiWWW 2024 · 被引用 32 次
- Exploiting Simulated User Feedback for Conversational Search: Ranking, Rewriting, and BeyondPaul Owoicho, Ivan Sekulic, Mohammad Aliannejadi, Jeffrey Dalton 等SIGIR 2023 · 被引用 31 次
- Knowledge-enhanced Mixed-initiative Dialogue System for Emotional Support ConversationsYang Deng, Wenxuan Zhang, Yifei Yuan, Wai LamACL 2023 · 被引用 31 次
- Structured and Natural Responses Co-generation for Conversational SearchChenchen Ye, Lizi Liao, Fuli Feng, Wei Ji 等SIGIR 2022 · 被引用 21 次
相关 Paper
- A LLM-based Controllable, Scalable, Human-Involved User Simulator Framework for Conversational Recommender SystemsLixi Zhu, Xiaowen Huang, Jitao SangWWW 2025 · 被引用 18 次
- Do Simulated Users Need to Remember? Analyzing the Impact of Memory Models in Conversational Search EvaluationNailia Mirzakhmedova, Marcel Gohsen, Johannes Kiesel, Matthias Hagen 等SIGIR 2026
- Analyzing and Simulating User Utterance Reformulation in Conversational Recommender SystemsShuo Zhang, Mu-Chun Wang, Krisztian BalogSIGIR 2022 · 被引用 18 次
- Task-Aware Automated User Profile Generation for Recommendation Simulation Using Large Language ModelsXinye Wanyan, Chenglong Ma, Danula Hettiachchi, Ziqi Xu 等SIGIR 2026
- Search-Based Interaction For Conversation Recommendation via Generative Reward Model Based Simulated UserXiaolei Wang, Chunxuan Xia, Junyi Li, Fanzhe Meng 等SIGIR 2025 · 被引用 1 次
