Gender Artifacts in Visual Datasets
Nicole Meister, Dora Zhao, Angelina Wang, Vikram V. Ramaswamy, Ruth Fong, Olga Russakovsky
Abstract
Gender biases are known to exist within large-scale visual datasets and can be reflected or even amplified in downstream models. Many prior works have proposed methods for mitigating gender biases, often by attempting to remove gender expression information from images. To understand the feasibility and practicality of these approaches, we investigate what "gender artifacts" exist in large-scale visual datasets. We define a "gender artifact" as a visual cue correlated with gender, focusing specifically on cues that are learnable by a modern image classifier and have an interpretable human corollary. Through our analyses, we find that gender artifacts are ubiquitous in the COCO and OpenImages datasets, occurring everywhere from low-level information (e.g., the mean value of the color channels) to higher-level image composition (e.g., pose and location of people). Further, bias mitigation methods that attempt to remove gender actually remove more information from the scene than the person. Given the prevalence of gender artifacts, we claim that attempts to remove these artifacts from such datasets are largely infeasible as certain removed artifacts may be necessary for the downstream task of object recognition. Instead, the responsibility lies with researchers and practitioners to be aware that the distribution of images within datasets is highly gendered and hence develop fairness-aware methods which are robust to these distributional shifts across groups.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 1d5fe28d-d375-4753-9242-bbc9e73d7239Cited by top-tier papers18
- FACET: Fairness in Computer Vision Evaluation BenchmarkLaura Gustafson, Chloé Rolland, Nikhila Ravi, Quentin Duval et al.ICCV 2023 · 74 citations
- Men Also Do Laundry: Multi-Attribute Bias AmplificationDora Zhao, Jerone Theodore Alexander Andrews, Alice XiangICML 2023 · 29 citations
- Understanding Bias in Large-Scale Visual DatasetsBoya Zeng, Yida Yin, Zhuang LiuNeurIPS 2024 · 23 citations
- MM-SHAP: A Performance-agnostic Metric for Measuring Multimodal Contributions in Vision and Language Models & TasksLetitia Parcalabescu, Anette FrankACL 2023 · 15 citations
- A Multidimensional Analysis of Social Biases in Vision TransformersJannik Brinkmann, Paul Swoboda, Christian BarteltICCV 2023 · 13 citations
Builds on12
- Balanced Datasets Are Not Enough: Estimating and Mitigating Gender Bias in Deep Image RepresentationsTianlu Wang, Jieyu Zhao, Mark Yatskar, Kai-Wei Chang et al.ICCV 2019 · 469 citations
- Noise or Signal: The Role of Image Backgrounds in Object RecognitionKai Yuanqing Xiao, Logan Engstrom, Andrew Ilyas, Aleksander MadryICLR 2021 · 451 citations
- Co-Designing Checklists to Understand Organizational Challenges and Opportunities around Fairness in AIMichael A. Madaio, Luke Stark, Jennifer Wortman Vaughan, Hanna M. WallachCHI 2020 · 428 citations
- How We've Taught Algorithms to See Identity: Constructing Race and Gender in Image Databases for Facial AnalysisMorgan Klaus Scheuerman, Kandrea Wade, Caitlin Lustig, Jed R. BrubakerCSCW 2020 · 198 citations
- Understanding and Evaluating Racial Biases in Image CaptioningDora Zhao, Angelina Wang, Olga RussakovskyICCV 2021 · 165 citations
Related papers
- Mitigating Gender Bias in Captioning SystemsRuixiang Tang, Mengnan Du, Yuening Li, Zirui Liu et al.WWW 2021 · 77 citations
- Bias in Gender Bias Benchmarks: How Spurious Features Distort EvaluationYusuke Hirota, Ryo Hachiuma, Boyi Li, Ximing Lu et al.ICCV 2025
- Common Sense Bias Modeling for Classification TasksMiao Zhang, Zee Fryer, Ben Colman, Ali Shahriyari et al.AAAI 2025
- Are Gender-Neutral Queries Really Gender-Neutral? Mitigating Gender Bias in Image SearchJialu Wang, Yang Liu, Xin Eric WangEMNLP 2021 · 44 citations
- MAVias: Mitigate any Visual BiasIoannis Sarridis, Christos Koutlis, Symeon Papadopoulos, Christos DiouICCV 2025 · 1 citation
