FairRec: Fairness Testing for Deep Recommender Systems
Huizhong Guo, Jinfeng Li, Jingyi Wang, Xiangyu Liu, Dongxia Wang, Zehong Hu, Rong Zhang, Hui Xue
Abstract
Deep learning-based recommender systems (DRSs) are increasingly and widely deployed in the industry, which brings significant convenience to people’s daily life in different ways. However, recommender systems are also shown to suffer from multiple issues, e.g., the echo chamber and the Matthew effect, of which the notation of “fairness” plays a core role. For instance, the system may be regarded as unfair to 1) a specific user, if the user gets worse recommendations than other users, or 2) an item (to recommend), if the item is much less likely to be exposed to the users than other items. While many fairness notations and corresponding fairness testing approaches have been developed for traditional deep classification models, they are essentially hardly applicable to DRSs. One major challenge is that there still lacks a systematic understanding and mapping between the existing fairness notations and the diverse testing requirements for deep recommender systems, not to mention further testing or debugging activities. To address the gap, we propose FairRec, a unified framework that supports fairness testing of DRSs from multiple customized perspectives, e.g., model utility, item diversity, item popularity, etc. We also propose a novel, efficient search-based testing approach to tackle the new challenge, i.e., double-ended discrete particle swarm optimization (DPSO) algorithm, to effectively search for hidden fairness issues in the form of certain disadvantaged groups from a vast number of candidate groups. Given the testing report, by adopting a simple re-ranking mitigation strategy on these identified disadvantaged groups, we show that the fairness of DRSs can be significantly improved. We conducted extensive experiments on multiple industry-level DRSs adopted by leading companies. The results confirm that FairRec is effective and efficient in identifying the deeply hidden fairness issues, e.g., achieving ∼95% testing accuracy with ∼half to 1/8 time.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers2
- A Large-Scale Empirical Study on Improving the Fairness of Image Classification ModelsJunjie Yang, Jiajun Jiang, Zeyu Sun, Junjie ChenISSTA 2024 · 4 citations
- Do Existing Testing Tools Really Uncover Gender Bias in Text-to-Image Models?Yunbo Lyu, Zhou Yang, Yuqing Niu, Jing Jiang et al.ACM MM 2025 · 2 citations
Builds on7
- User-oriented Fairness in RecommendationYunqi Li, Hanxiong Chen, Zuohui Fu, Yingqiang Ge et al.WWW 2021 · 293 citations
- White-box fairness testing through adversarial samplingPeixin Zhang, Jingyi Wang, Jun Sun, Guoliang Dong et al.ICSE 2020 · 127 citations
- Debiasing Career Recommendations with Neural Fair Collaborative FilteringRashidul Islam, Kamrun Naher Keya, Ziqian Zeng, Shimei Pan et al.WWW 2021 · 85 citations
- Popularity Bias in Dynamic RecommendationZiwei Zhu, Yun He, Xing Zhao, James CaverleeKDD 2021 · 77 citations
- Explanation-Guided Fairness Testing through Genetic AlgorithmMing Fan, Wenying Wei, Wuxia Jin, Zijiang Yang et al.ICSE 2022 · 45 citations
Related papers
- No Bias Left Behind: Fairness Testing for Deep Recommender Systems Targeting General Disadvantaged GroupsZhuo Wu, Zan Wang, Chuan Luo, Xiaoning Du et al.ISSTA 2025
- Preference Amplification in Recommender SystemsDimitris Kalimeris, Smriti Bhagat, Shankar Kalyanaraman, Udi WeinsbergKDD 2021 · 24 citations
- Diversity Matters: User-Centric Multi-Interest Learning for Conversational Movie RecommendationYongsen Zheng, Guohua Wang, Yang Liu, Liang LinACM MM 2024 · 5 citations
- HyCoRec: Hypergraph-Enhanced Multi-Preference Learning for Alleviating Matthew Effect in Conversational RecommendationYongsen Zheng, Ruilin Xu, Ziliang Chen, Guohua Wang et al.ACL 2024
- Mitigating Matthew Effect: Multi-Hypergraph Boosted Multi-Interest Self-Supervised Learning for Conversational RecommendationYongsen Zheng, Ruilin Xu, Guohua Wang, Liang Lin et al.EMNLP 2024 · 2 citations
