On Sampled Metrics for Item Recommendation
Walid Krichene, Steffen Rendle
摘要
The task of item recommendation requires ranking a large catalogue of items given a context. Item recommendation algorithms are evaluated using ranking metrics that depend on the positions of relevant items. To speed up the computation of metrics, recent work often uses sampled metrics where only a smaller set of random items and the relevant items are ranked. This paper investigates sampled metrics in more detail and shows that they are inconsistent with their exact version, in the sense that they do not persist relative statements, e.g., recommender A is better than B, not even in expectation. Moreover, the smaller the sampling size, the less difference there is between metrics, and for very small sampling size, all metrics collapse to the AUC metric. We show that it is possible to improve the quality of the sampled metrics by applying a correction, obtained by minimizing different criteria such as bias or mean squared error. We conclude with an empirical evaluation of the naive sampled metrics and their corrected variants. To summarize, our work suggests that sampling should be avoided for metric calculation, however if an experimental study needs to sample, the proposed corrections can improve the quality of the estimate.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper88
- Contrastive Learning for Sequential RecommendationXu Xie, Fei Sun, Zhaoyang Liu, Shiwen Wu 等ICDE 2022 · 被引用 674 次
- Learning Intents behind Interactions with Knowledge Graph for RecommendationXiang Wang, Tinglin Huang, Dingxian Wang, Yancheng Yuan 等WWW 2021 · 被引用 584 次
- Intent Contrastive Learning for Sequential RecommendationYongjun Chen, Zhiwei Liu, Jia Li, Julian J. McAuley 等WWW 2022 · 被引用 429 次
- Sequential Recommendation via Stochastic Self-AttentionZiwei Fan, Zhiwei Liu, Yu Wang, Alice Wang 等WWW 2022 · 被引用 203 次
- MixGCF: An Improved Training Method for Graph Neural Network-based Recommender SystemsTinglin Huang, Yuxiao Dong, Ming Ding, Zhen Yang 等KDD 2021 · 被引用 190 次
相关 Paper
- Towards Reliable Item Sampling for Recommendation EvaluationDong Li, Ruoming Jin, Zhenming Liu, Bin Ren 等AAAI 2023 · 被引用 11 次
- On Estimating Recommendation Evaluation Metrics under SamplingRuoming Jin, Dong Li, Benjamin Mudrak, Jing Gao 等AAAI 2021 · 被引用 16 次
- Lower-Left Partial AUC: An Effective and Efficient Optimization Metric for RecommendationWentao Shi, Chenxu Wang, Fuli Feng, Yang Zhang 等WWW 2024 · 被引用 13 次
- Agreement and Disagreement between True and False-Positive Metrics in Recommender Systems EvaluationElisa Mena-Maldonado, Rocío Cañamares, Pablo Castells, Yongli Ren 等SIGIR 2020 · 被引用 16 次
- On Sampling Top-K Recommendation EvaluationDong Li, Ruoming Jin, Jing Gao, Zhi LiuKDD 2020 · 被引用 43 次
