Association of Objects May Engender Stereotypes: Mitigating Association-Engendered Stereotypes in Text-to-Image Generation
Junlei Zhou, Jiashi Gao, Xiangyu Zhao, Xin Yao, Xuetao Wei
摘要
Text-to-Image (T2I) has witnessed significant advancements, demonstrating superior performance for various generative tasks. However, the presence of stereotypes in T2I introduces harmful biases that require urgent attention as the T2I technology becomes more prominent. Previous work for stereotype mitigation mainly concentrated on mitigating stereotypes engendered with individual objects within images, which failed to address stereotypes engendered by the association of multiple objects, referred to as Association-Engendered Stereotypes . For example, mentioning “black people” and “houses” separately in prompts may not exhibit stereotypes. Nevertheless, when these two objects are associated in prompts, the association of “black people” with “poorer houses” becomes more pronounced. To tackle this issue, we propose a novel framework, MAS , to Mitigate Association-engendered Stereotypes. This framework models the stereotype problem as a probability distribution alignment problem, aiming to align the stereotype probability distribution of the generated image with the stereotype-free distribution. The MAS framework primarily consists of the Prompt-Image-Stereotype CLIP ( PIS CLIP ) and Sensitive Transformer . The PIS CLIP learns the association between prompts, images, and stereotypes, which can establish the mapping of prompts to stereotypes. The Sensitive Transformer produces the sensitive constraints, which guide the stereotyped image distribution to align with the stereotype-free probability distribution. Moreover, recognizing that existing metrics are insufficient for accurately evaluating association-engendered stereotypes, we propose a novel metric, Stereotype-Distribution-Total-Variation ( SDTV ), to evaluate stereotypes in T2I. Comprehensive experiments demonstrate that our framework effectively mitigates association-engendered stereotypes
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- Fair Generation without Unfair Distortions: Debiasing Text-To-Image Generation with Entanglement-Free AttentionJeonghoon Park, Juyoung Lee, Chaeyeon Chung, Jaeseong Lee 等ICCV 2025 · 被引用 1 次
- Causally-Grounded Dual-Path Attention Intervention for Object Hallucination Mitigation in LVLMsLiu Yu, Zhonghao Chen, Ping Kuang, Zhikun Feng 等AAAI 2026
它引用的顶会 Paper15
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- Photorealistic Text-to-Image Diffusion Models with Deep Language UnderstandingChitwan Saharia, William Chan, Saurabh Saxena, Lala Li 等NeurIPS 2022 · 被引用 8,965 次
- Adding Conditional Control to Text-to-Image Diffusion ModelsLvmin Zhang, Anyi Rao, Maneesh AgrawalaICCV 2023 · 被引用 6,759 次
- GLIDE: Towards Photorealistic Image Generation and Editing with Text-Guided Diffusion ModelsAlexander Quinn Nichol, Prafulla Dhariwal, Aditya Ramesh, Pranav Shyam 等ICML 2022 · 被引用 4,691 次
相关 Paper
- Multi-Group Proportional Representations for Text-to-Image ModelsSangwon Jung, Alex Oesterling, Claudio Mayrink Verdun, Sajani Vithana 等CVPR 2025
- Mitigating Stereotypes in Text-to-Image Generation: A Novel Perspective of Selective Neural SuppressionJunlei Zhou, Jiashi Gao, Xinwei Guo, Haiyan Wu 等ACM MM 2025 · 被引用 1 次
- AutoDebias: An Automated Framework for Detecting and Mitigating Backdoor Biases in Text-to-Image ModelsHongyi Cai, Mohammad Mahdinur Rahman, MingKang Dong, Muxin Pu 等CVPR 2026
- The Male CEO and the Female Assistant: Evaluation and Mitigation of Gender Biases in Text-To-Image Generation of Dual SubjectsYixin Wan, Kai-Wei ChangACL 2025 · 被引用 11 次
- OASIS Uncovers: High-Quality T2I Models, Same Old StereotypesSepehr Dehdashtian, Gautam Sreekumar, Vishnu BoddetiICLR 2025
