To "See" is to Stereotype: Image Tagging Algorithms, Gender Recognition, and the Accuracy-Fairness Trade-off
Pinar Barlas, Kyriakos Kyriakou, Olivia Guest, Styliani Kleanthous, Jahna Otterbacher
摘要
Machine-learned computer vision algorithms for tagging images are increasingly used by developers and researchers, having become popularized as easy-to-use "cognitive services." Yet these tools struggle with gender recognition, particularly when processing images of women, people of color and non-binary individuals. Socio-technical researchers have cited data bias as a key problem; training datasets often over-represent images of people and contexts that convey social stereotypes. The social psychology literature explains that people learn social stereotypes, in part, by observing others in particular roles and contexts, and can inadvertently learn to associate gender with scenes, occupations and activities. Thus, we study the extent to which image tagging algorithms mimic this phenomenon. We design a controlled experiment, to examine the interdependence between algorithmic recognition of context and the depicted person's gender. In the spirit of auditing to understand machine behaviors, we create a highly controlled dataset of people images, imposed on gender-stereotyped backgrounds. Our methodology is reproducible and our code publicly available. Evaluating five proprietary algorithms, we find that in three, gender inference is hindered when a background is introduced. Of the two that "see" both backgrounds and gender, it is the one whose output is most consistent with human stereotyping processes that is superior in recognizing gender. We discuss the accuracy--fairness trade-off, as well as the importance of auditing black boxes in better understanding this double-edged sword.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- Cruising Queer HCI on the DL: A Literature Review of LGBTQ+ People in HCIJordan Taylor, Ellen Simpson, Anh-Ton Tran, Jed R. Brubaker 等CHI 2024 · 被引用 55 次
- RES: A Robust Framework for Guiding Visual ExplanationYuyang Gao, Tong Steven Sun, Guangji Bai, Siyi Gu 等KDD 2022 · 被引用 29 次
- MEDebiaser: A Human-AI Feedback System for Mitigating Bias in Multi-label Medical Image ClassificationShaohan Shi, Yuheng Shao, Haoran Jiang, Yunjie Yao 等UIST 2025
- Rethinking Pareto Frontier: On the Optimal Trade-offs in Fair ClassificationJunyi Chai, Shenyu Lu, Xiaoqian WangICLR 2026
相关 Paper
- Understanding and Evaluating Racial Biases in Image CaptioningDora Zhao, Angelina Wang, Olga RussakovskyICCV 2021 · 被引用 165 次
- Balanced Datasets Are Not Enough: Estimating and Mitigating Gender Bias in Deep Image RepresentationsTianlu Wang, Jieyu Zhao, Mark Yatskar, Kai-Wei Chang 等ICCV 2019 · 被引用 469 次
- Gender Artifacts in Visual DatasetsNicole Meister, Dora Zhao, Angelina Wang, Vikram V. Ramaswamy 等ICCV 2023 · 被引用 37 次
- EuroGEST: Investigating gender stereotypes in multilingual language modelsJacqueline Rowe, Mateusz Klimaszewski, Liane Guillou, Shannon Vallor 等EMNLP 2025
- Sensemaking in User-Driven Algorithm Auditing: A Case Study on Gender Bias in an Image Captioning ModelBehnoosh Mohammadzadeh, Jules Françoise, Michèle Gouiffès, Baptiste CaramiauxCHI 2026 · 被引用 1 次
