New Job, New Gender? Measuring the Social Bias in Image Generation Models
Wenxuan Wang, Haonan Bai, Jen-tse Huang, Yuxuan Wan, Youliang Yuan, Haoyi Qiu, Nanyun Peng, Michael R. Lyu
Abstract
Image generation models can generate or edit images from a given text. Recent advancements in image generation technology, exemplified by DALL-E and Midjourney, have been groundbreaking. These advanced models, despite their impressive capabilities, are often trained on massive Internet datasets, making them susceptible to generating content that perpetuates social stereotypes and biases, which can lead to severe consequences. Prior research on assessing bias within image generation models suffers from several shortcomings, including limited accuracy, reliance on extensive human labor, and lack of comprehensive analysis. In this paper, we propose BiasPainter, a novel evaluation framework that can accurately, automatically and comprehensively trigger social bias in image generation models. BiasPainter uses a diverse range of seed images of individuals and prompts the image generation models to edit these images using gender, race, and age-neutral queries. These queries span 62 professions, 39 activities, 57 types of objects, and 70 personality traits. The framework then compares the edited images to the original seed images, focusing on the significant changes related to gender, race, and age. BiasPainter adopts a key insight that these characteristics should not be modified when subjected to neutral prompts. Built upon this design, BiasPainter can trigger the social bias and evaluate the fairness of image generation models. We use BiasPainter to evaluate six widely-used image generation models, such as stable diffusion and Midjourney. Experimental results show that BiasPainter can successfully trigger social bias in image generation models. According to our human evaluation, BiasPainter can achieve 90.8% accuracy on automatic bias detection, which is significantly higher than the results reported in previous work.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext ef880283-bcbe-44af-bb7c-5efae341333eCited by top-tier papers6
- Ferrari: Federated Feature Unlearning via Optimizing Feature SensitivityHanlin Gu, WinKent Ong, Chee Seng Chan, Lixin FanNeurIPS 2024 · 29 citations
- FairImagen: Post-Processing for Bias Mitigation in Text-to-Image ModelsZihao Fu, Ryan Brown, Shun Shao, Kai Rawal et al.NeurIPS 2025 · 5 citations
- Fairness Mediator: Neutralize Stereotype Associations to Mitigate Bias in Large Language ModelsYisong Xiao, Aishan Liu, Siyuan Liang, Xianglong Liu et al.ISSTA 2025 · 2 citations
- Do Existing Testing Tools Really Uncover Gender Bias in Text-to-Image Models?Yunbo Lyu, Zhou Yang, Yuqing Niu, Jing Jiang et al.ACM MM 2025 · 2 citations
- AI Sees Your Location - But With A Bias Toward The Wealthy WorldJingyuan Huang, Jen-tse Huang, Ziyi Liu, Xiaoyuan Liu et al.EMNLP 2025
Builds on10
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- SDXL: Improving Latent Diffusion Models for High-Resolution Image SynthesisDustin Podell, Zion English, Kyle Lacey, Andreas Blattmann et al.ICLR 2024 · 4,569 citations
- Towards Understanding and Mitigating Social Biases in Language ModelsPaul Pu Liang, Chiyu Wu, Louis-Philippe Morency, Ruslan SalakhutdinovICML 2021 · 495 citations
- You Keep Using That Word: Ways of Thinking about Gender in Computing ResearchOs Keyes, Chandler May, Annabelle CarrellCSCW 2021 · 33 citations
Related papers
- Exposing Hidden Biases in Text-to-Image Models via Automated Prompt SearchManos Plitsis, Giorgos Bouritsas, Vassilis Katsouros, Yannis PanagakisICML 2026
- SustainDiffusion: Optimising the Social and Environmental Sustainability of Stable Diffusion ModelsGiordano d’Aloisio, Tosin Fadahunsi, Jay Choy, Rebecca Moussa et al.ICSE 2026
- Image-Perfect Imperfections: Safety, Bias, and Authenticity in the Shadow of Text-To-Image Model EvolutionYixin Wu, Yun Shen, Michael Backes, Yang ZhangCCS 2024 · 3 citations
- DALL-EVAL: Probing the Reasoning Skills and Social Biases of Text-to-Image Generation ModelsJaemin Cho, Abhay Zala, Mohit BansalICCV 2023 · 283 citations
- Mitigating Social Biases in Text-to-Image Diffusion Models via Linguistic-Aligned Attention GuidanceYue Jiang, Yueming Lyu, Ziwen He, Bo Peng et al.ACM MM 2024 · 4 citations
