Societal Biases in Retrieved Contents: Measurement Framework and Adversarial Mitigation of BERT Rankers
Navid Rekabsaz, Simone Kopeinik, Markus Schedl
Abstract
Societal biases resonate in the retrieved contents of information retrieval (IR) systems, resulting in reinforcing existing stereotypes. Approaching this issue requires established measures of fairness in respect to the representation of various social groups in retrieval results, as well as methods to mitigate such biases, particularly in the light of the advances in deep ranking models. In this work, we first provide a novel framework to measure the fairness in the retrieved text contents of ranking models. Introducing a ranker-agnostic measurement, the framework also enables the disentanglement of the effect on fairness of collection from that of rankers. To mitigate these biases, we propose AdvBert, a ranking model achieved by adapting adversarial bias mitigation for IR, which jointly learns to predict relevance and remove protected attributes. We conduct experiments on two passage retrieval collections (MSMARCO Passage Re-ranking and TREC Deep Learning 2019 Passage Re-ranking), which we extend by fairness annotations of a selected subset of queries regarding gender attributes. Our results on the MSMARCO benchmark show that, (1) all ranking models are less fair in comparison with ranker-agnostic baselines, and (2) the fairness of Bert rankers significantly improves when using the proposed AdvBert models. Lastly, we investigate the trade-off between fairness and utility, showing that we can maintain the significant improvements in fairness without any significant loss in utility.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 0bd3540a-199b-4811-b8f2-001d6be4db58Cited by top-tier papers4
- Joint Multisided Exposure Fairness for RecommendationHaolun Wu, Bhaskar Mitra, Chen Ma, Fernando Diaz et al.SIGIR 2022 · 48 citations
- A Multidimensional Analysis of Social Biases in Vision TransformersJannik Brinkmann, Paul Swoboda, Christian BarteltICCV 2023 · 13 citations
- Algorithmic Vibe in Information RetrievalAli Montazeralghaem, Nick Craswell, Ryen W. White, Ahmed Hassan Awadallah et al.WWW 2023
- With Argus Eyes: Assessing Retrieval Gaps via Uncertainty Scoring to Detect and Remedy Retrieval Blind SpotsZeinab Taghavi, Ali Modarressi, Hinrich Schuetze, Andreas MarfurtICML 2026
Builds on5
- Approximate Nearest Neighbor Negative Contrastive Learning for Dense Text RetrievalLee Xiong, Chenyan Xiong, Ye Li, Kwok-Fung Tang et al.ICLR 2021 · 1,547 citations
- ColBERT: Efficient and Effective Passage Search via Contextualized Late Interaction over BERTOmar Khattab, Matei ZahariaSIGIR 2020 · 1,246 citations
- Controlling Fairness and Bias in Dynamic Learning-to-RankMarco Morik, Ashudeep Singh, Jessica Hong, Thorsten JoachimsSIGIR 2020 · 205 citations
- Language (Technology) is Power: A Critical Survey of "Bias" in NLPSu Lin Blodgett, Solon Barocas, Hal Daumé III, Hanna M. WallachACL 2020 · 68 citations
- Efficient Document Re-Ranking for Transformers by Precomputing Term RepresentationsSean MacAvaney, Franco Maria Nardini, Raffaele Perego, Nicola Tonellotto et al.SIGIR 2020 · 62 citations
Related papers
- Fairness-Aware Exposure Allocation via Adaptive RerankingThomas Jänich, Graham McDonald, Iadh OunisSIGIR 2024 · 14 citations
- Fair Ranking with Noisy Protected AttributesAnay Mehrotra, Nisheeth K. VishnoiNeurIPS 2022 · 24 citations
- A Unified Pretraining Framework for Passage Ranking and ExpansionMing Yan, Chenliang Li, Bin Bi, Wei Wang et al.AAAI 2021 · 14 citations
- RedditBias: A Real-World Resource for Bias Evaluation and Debiasing of Conversational Language ModelsSoumya Barikeri, Anne Lauscher, Ivan Vulic, Goran GlavasACL 2021
- Mitigating Test-Time Bias for Fair Image RetrievalFanjie Kong, Shuai Yuan, Weituo Hao, Ricardo HenaoNeurIPS 2023 · 30 citations
