On Sampled Metrics for Item Recommendation
Walid Krichene, Steffen Rendle
Abstract
The task of item recommendation requires ranking a large catalogue of items given a context. Item recommendation algorithms are evaluated using ranking metrics that depend on the positions of relevant items. To speed up the computation of metrics, recent work often uses sampled metrics where only a smaller set of random items and the relevant items are ranked. This paper investigates sampled metrics in more detail and shows that they are inconsistent with their exact version, in the sense that they do not persist relative statements, e.g., recommender A is better than B, not even in expectation. Moreover, the smaller the sampling size, the less difference there is between metrics, and for very small sampling size, all metrics collapse to the AUC metric. We show that it is possible to improve the quality of the sampled metrics by applying a correction, obtained by minimizing different criteria such as bias or mean squared error. We conclude with an empirical evaluation of the naive sampled metrics and their corrected variants. To summarize, our work suggests that sampling should be avoided for metric calculation, however if an experimental study needs to sample, the proposed corrections can improve the quality of the estimate.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext c2585457-f512-4f14-93fa-a4b5ed5369e0Cited by top-tier papers88
- Contrastive Learning for Sequential RecommendationXu Xie, Fei Sun, Zhaoyang Liu, Shiwen Wu et al.ICDE 2022 · 674 citations
- Learning Intents behind Interactions with Knowledge Graph for RecommendationXiang Wang, Tinglin Huang, Dingxian Wang, Yancheng Yuan et al.WWW 2021 · 584 citations
- Intent Contrastive Learning for Sequential RecommendationYongjun Chen, Zhiwei Liu, Jia Li, Julian J. McAuley et al.WWW 2022 · 429 citations
- Sequential Recommendation via Stochastic Self-AttentionZiwei Fan, Zhiwei Liu, Yu Wang, Alice Wang et al.WWW 2022 · 203 citations
- MixGCF: An Improved Training Method for Graph Neural Network-based Recommender SystemsTinglin Huang, Yuxiao Dong, Ming Ding, Zhen Yang et al.KDD 2021 · 190 citations
Related papers
- Towards Reliable Item Sampling for Recommendation EvaluationDong Li, Ruoming Jin, Zhenming Liu, Bin Ren et al.AAAI 2023 · 11 citations
- On Estimating Recommendation Evaluation Metrics under SamplingRuoming Jin, Dong Li, Benjamin Mudrak, Jing Gao et al.AAAI 2021 · 16 citations
- Lower-Left Partial AUC: An Effective and Efficient Optimization Metric for RecommendationWentao Shi, Chenxu Wang, Fuli Feng, Yang Zhang et al.WWW 2024 · 13 citations
- Agreement and Disagreement between True and False-Positive Metrics in Recommender Systems EvaluationElisa Mena-Maldonado, Rocío Cañamares, Pablo Castells, Yongli Ren et al.SIGIR 2020 · 16 citations
- On Sampling Top-K Recommendation EvaluationDong Li, Ruoming Jin, Jing Gao, Zhi LiuKDD 2020 · 43 citations
