Understanding and Evaluating Racial Biases in Image Captioning
Dora Zhao, Angelina Wang, Olga Russakovsky
Abstract
Image captioning is an important task for benchmarking visual reasoning and for enabling accessibility for people with vision impairments. However, as in many machine learning settings, social biases can influence image captioning in undesirable ways. In this work, we study bias propagation pathways within image captioning, focusing specifically on the COCO dataset. Prior work has analyzed gender bias in captions using automatically-derived gender labels; here we examine racial and intersectional biases using manual annotations. Our first contribution is in annotating the perceived gender and skin color of 28,315 of the depicted people after obtaining IRB approval. Using these annotations, we compare racial biases present in both manual and automatically-generated image captions. We demonstrate differences in caption performance, sentiment, and word choice between images of lighter versus darker-skinned people. Further, we find the magnitude of these differences to be greater in modern captioning systems compared to older ones, thus leading to concerns that without proper consideration and mitigation these differences will only become increasingly prevalent. Code and data is available at https://princetonvisualai. github.io/imagecaptioning-bias/ .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 4a9a1b36-c0cf-41ee-810e-954b4d8518a8Cited by top-tier papers40
- Flamingo: a Visual Language Model for Few-Shot LearningJean-Baptiste Alayrac, Jeff Donahue, Pauline Luc, Antoine Miech et al.NeurIPS 2022 · 6,707 citations
- DALL-EVAL: Probing the Reasoning Skills and Social Biases of Text-to-Image Generation ModelsJaemin Cho, Abhay Zala, Mohit BansalICCV 2023 · 283 citations
- FACET: Fairness in Computer Vision Evaluation BenchmarkLaura Gustafson, Chloé Rolland, Nikhila Ravi, Quentin Duval et al.ICCV 2023 · 74 citations
- On Learning Fairness and Accuracy on Multiple SubgroupsChangjian Shui, Gezheng Xu, Qi Chen, Jiaqi Li et al.NeurIPS 2022 · 58 citations
- Inspecting the Geographical Representativeness of Images from Text-to-Image ModelsAbhipsa Basu, R. Venkatesh Babu, Danish PruthiICCV 2023 · 54 citations
Builds on10
- Attention on Attention for Image CaptioningLun Huang, Wenmin Wang, Jie Chen, Xiaoyong WeiICCV 2019 · 992 citations
- Balanced Datasets Are Not Enough: Estimating and Mitigating Gender Bias in Deep Image RepresentationsTianlu Wang, Jieyu Zhao, Mark Yatskar, Kai-Wei Chang et al.ICCV 2019 · 469 citations
- How We've Taught Algorithms to See Identity: Constructing Race and Gender in Image Databases for Facial AnalysisMorgan Klaus Scheuerman, Kandrea Wade, Caitlin Lustig, Jed R. BrubakerCSCW 2020 · 198 citations
- A Study of Face Obfuscation in ImageNetKaiyu Yang, Jacqueline H. Yau, Li Fei-Fei, Jia Deng et al.ICML 2022 · 163 citations
- Fair Generative Modeling via Weak SupervisionKristy Choi, Aditya Grover, Trisha Singh, Rui Shu et al.ICML 2020 · 160 citations
Related papers
- Uncurated Image-Text Datasets: Shedding Light on Demographic BiasNoa Garcia, Yusuke Hirota, Yankun Wu, Yuta NakashimaCVPR 2023
- Mitigating Gender Bias in Captioning SystemsRuixiang Tang, Mengnan Du, Yuening Li, Zirui Liu et al.WWW 2021 · 77 citations
- Men Also Do Laundry: Multi-Attribute Bias AmplificationDora Zhao, Jerone Theodore Alexander Andrews, Alice XiangICML 2023 · 29 citations
- Gender Biases in Automatic Evaluation Metrics for Image CaptioningHaoyi Qiu, Zi-Yi Dou, Tianlu Wang, Asli Celikyilmaz et al.EMNLP 2023 · 6 citations
- Quantifying Societal Bias Amplification in Image CaptioningYusuke Hirota, Yuta Nakashima, Noa GarciaCVPR 2022 · 45 citations
