Standing in Your Shoes: External Assessments for Personalized Recommender Systems
Hongyu Lu, Weizhi Ma, Min Zhang, Maarten de Rijke, Yiqun Liu, Shaoping Ma
Abstract
The evaluation of recommender systems relies on user preference data, which is difficult to acquire directly because of its subjective nature. Current recommender systems widely utilize users' historical interactions as implicit or explicit feedback, but such data usually suffers from various types of bias. Little work has been done on collecting and understanding user's personal preferences via third-party annotations. External assessments, that is, annotations made by assessors who are not the systems' users, have been widely used in information search scenarios. Is it possible to use external assessments to construct user preference labels? This paper presents the first attempt to incorporate external assessments into preference labeling and recommendation evaluation. The aim is to verify the possibility and reliability of external assessments for personalized recommender systems. We collect both users' real preferences and assessors' estimated preferences through a multi-role, multi-session user study. By investigating the inter-assessor agreement and user-assessor consistency, we demonstrate the reasonable stability and high accuracy of external preference assessments. Furthermore, we investigate the usage of external assessments in system evaluation. A higher degree of consistency with users' online feedback is observed, even better than traditional history-based online evaluation. Our findings show that external assessments can be used for assessing user preference labels and evaluating systems in personalized recommendation scenarios.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext dc92dc6c-8ec0-4581-b740-6663ca5bcef4Related papers
- Evaluating Conversational Recommender Systems via User SimulationShuo Zhang, Krisztian BalogKDD 2020 · 80 citations
- Context-based Collective Preference Aggregation for Prioritizing Crowd Opinions in Social Decision-makingJiyi LiWWW 2022 · 3 citations
- Quantifying Assimilate-Contrast Effects in Online Rating Systems: Modeling, Analysis and ApplicationMingze Zhong, Hong Xie, Qingsheng ZhuKDD 2021 · 1 citation
- Large Language Models can Accurately Predict Searcher PreferencesPaul Thomas, Seth Spielman, Nick Craswell, Bhaskar MitraSIGIR 2024 · 153 citations
- Studying the Effects of Cognitive Biases in Evaluation of Conversational AgentsSashank Santhanam, Alireza Karduni, Samira ShaikhCHI 2020 · 18 citations
