On Fairness of Unified Multimodal Large Language Model for Image Generation
Ming Liu, Hao Chen, Jindong Wang, Liwen Wang, Bhiksha Raj, Wensheng Zhang
摘要
Unified multimodal large language models (U-MLLMs) have demonstrated impressive performance in end-to-end visual understanding and generation tasks. However, compared to generation-only systems (e.g., Stable Diffusion), the unified architecture of U-MLLMs introduces new risks of propagating demographic stereotypes. In this paper, we benchmark several state-of-the-art U-MLLMs and show that they exhibit significant gender and race biases in the generated outputs. To diagnose the source of these biases, we propose a locate-then-fix framework: we first audit the vision and language components -using techniques such as linear probing and controlled generation -and find that the language model appears to be a primary origin of the observed generative bias. Moreover, we observe a "partial alignment" phenomenon, where the U-MLLMs exhibit less bias in understanding tasks yet produce substantially biased images. To address this, we introduce a novel balanced preference loss that enforces uniform generation probabilities across demographics by leveraging a synthetically balanced dataset. Extensive experiments show that our approach significantly reduces demographic bias while preserving semantic fidelity and image quality. Our findings underscore the need for targeted debiasing strategies in unified multimodal systems and introduce a practical approach to mitigate biases.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- Doctor Approved: Generating Medically Accurate Skin Disease Images through AI-Expert FeedbackJanet Wang, Yunbei Zhang, Zhengming Ding, Jihun HammNeurIPS 2025 · 被引用 13 次
- When Understanding Becomes a Risk: Authenticity and Safety Risks in the Emerging Image Generation ParadigmYe Leng, Junjie Chu, Mingjie Li, Chenhao Lin 等CVPR 2026 · 被引用 3 次
- Fair in Mind, Fair in Action? A Synchronous Benchmark for Understanding and Generation in UMLLMsYiran Zhao, Lu Zhou, Xiaogang Xu, Zhe Liu 等ICLR 2026 · 被引用 1 次
它引用的顶会 Paper27
- Training language models to follow instructions with human feedbackLong Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida 等NeurIPS 2022 · 被引用 24,707 次
- LoRA: Low-Rank Adaptation of Large Language ModelsEdward J. Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu 等ICLR 2022 · 被引用 18,833 次
- Visual Instruction TuningHaotian Liu, Chunyuan Li, Qingyang Wu, Yong Jae LeeNeurIPS 2023 · 被引用 11,349 次
- Direct Preference Optimization: Your Language Model is Secretly a Reward ModelRafael Rafailov, Archit Sharma, Eric Mitchell, Christopher D. Manning 等NeurIPS 2023 · 被引用 10,924 次
- Exploring CLIP for Assessing the Look and Feel of ImagesJianyi Wang, Kelvin C. K. Chan, Chen Change LoyAAAI 2023 · 被引用 1,208 次
相关 Paper
- Mitigating Social Biases in Text-to-Image Diffusion Models via Linguistic-Aligned Attention GuidanceYue Jiang, Yueming Lyu, Ziwen He, Bo Peng 等ACM MM 2024 · 被引用 4 次
- Fine-tuning Bias Neurons for Fair Text-to-Image GenerationFan Qi, Zhan Wang, Changsheng Xu, Huaiwen ZhangACM MM 2025
- Person-Centric Annotations of LAION-400M: Auditing Bias and Its Transfer to ModelsLeander Girrbach, Stephan Alaniz, Genevieve Smith, Trevor Darrell 等ICLR 2026 · 被引用 5 次
- Debiasing Multimodal Large Language Models via Penalization of Language PriorsYifan Zhang, Yang Shi, Weichen Yu, Qingsong Wen 等ACM MM 2025 · 被引用 6 次
- Social Debiasing for Fair Multi-Modal LLMsHarry Cheng, Yangyang Guo, Qing Guo, Ming-Hsuan Yang 等ICCV 2025 · 被引用 1 次
