Can We Trust Recommender System Fairness Evaluation? The Role of Fairness and Relevance
Theresia Veronika Rampisela, Tuukka Ruotsalo, Maria Maistro, Christina Lioma
摘要
Relevance and fairness are two major objectives of recommender systems (RSs). Recent work proposes measures of RS fairness that are either independent from relevance (fairness-only) or conditioned on relevance (joint measures). While fairness-only measures have been studied extensively, we look into whether joint measures can be trusted. We collect all joint evaluation measures of RS relevance and fairness, and ask: How much do they agree with each other? To what extent do they agree with relevance/fairness measures? How sensitive are they to changes in rank position, or to increasingly fair and relevant recommendations? We eempirically study for the first time the behaviour of these measures across 4 real-world datasets and 4 recommenders. We find that most of these measures: i) correlate weakly with one another and even contradict each other at times; ii) are less sensitive to rank position changes than relevance- and fairness-only measures, meaning that they are less granular than traditional RS measures; and iii) tend to compress scores at the low end of their range, meaning that they are not very expressive. We counter the above limitations with a set of guidelines on the appropriate usage of such measures, i.e., they should be used with caution due to their tendency to contradict each other and of having a very small empirical range.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- Joint Evaluation of Fairness and Relevance in Recommender Systems with Pareto FrontierTheresia Veronika Rampisela, Tuukka Ruotsalo, Maria Maistro, Christina LiomaWWW 2025 · 被引用 3 次
- Explaining Rankings with Hidden Group BonusesAlvin Hong Yao Yan, Suraj Shetiya, Sujoy Bhore, Priyanka Golia 等KDD 2026
它引用的顶会 Paper10
- Improving Graph Collaborative Filtering with Neighborhood-enriched Contrastive LearningZihan Lin, Changxin Tian, Yupeng Hou, Wayne Xin ZhaoWWW 2022 · 被引用 606 次
- FairRec: Two-Sided Fairness for Personalized Recommendations in Two-Sided PlatformsGourab K. Patro, Arpita Biswas, Niloy Ganguly, Krishna P. Gummadi 等WWW 2020 · 被引用 268 次
- Measuring Fairness in Ranked Results: An Analytical and Empirical ComparisonAmifa Raj, Michael D. EkstrandSIGIR 2022 · 被引用 74 次
- Computationally Efficient Optimization of Plackett-Luce Ranking Models for Relevance and FairnessHarrie OosterhuisSIGIR 2021 · 被引用 68 次
- Joint Multisided Exposure Fairness for RecommendationHaolun Wu, Bhaskar Mitra, Chen Ma, Fernando Diaz 等SIGIR 2022 · 被引用 48 次
相关 Paper
- User-item fairness tradeoffs in recommendationsSophie Greenwood, Sudalakshmee Chiniah, Nikhil GargNeurIPS 2024 · 被引用 15 次
- Fair Representation Learning for Recommendation: A Mutual Information PerspectiveChen Zhao, Le Wu, Pengyang Shao, Kun Zhang 等AAAI 2023 · 被引用 37 次
- CPFair: Personalized Consumer and Producer Fairness Re-ranking for Recommender SystemsMohammadmehdi Naghiaei, Hossein A. Rahmani, Yashar DeldjooSIGIR 2022 · 被引用 117 次
- Fairness among New Items in Cold Start Recommender SystemsZiwei Zhu, Jingu Kim, Trung Nguyen, Aish Fenton 等SIGIR 2021 · 被引用 74 次
- Agreement and Disagreement between True and False-Positive Metrics in Recommender Systems EvaluationElisa Mena-Maldonado, Rocío Cañamares, Pablo Castells, Yongli Ren 等SIGIR 2020 · 被引用 16 次
