Plug-and-Play Interpretable Responsible Text-to-Image Generation via Dual-Space Multi-facet Concept Control
Basim Azam, Naveed Akhtar
Abstract
Figure 1 . Our unique plug-and-play interpretable approach simultaneously controls a range of concepts for responsible and fair image generation with text-to-image pipelines. Our method enables control over both text embedding space and latent diffusion space. Shown examples compare unfair and unsafe generation for the given prompts by the Stable Diffusion (top), along with their responsible counterparts resulting from our approach influencing the text encoder (middle) and the diffusion model (bottom). We control individual concepts for diverse/safe generation in these examples while modeling them in continuous composite responsible semantic spaces.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext c80d154d-6541-487b-b5e0-4bd74ab37f49Cited by top-tier papers1
Ask how each one uses itBuilds on25
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- Photorealistic Text-to-Image Diffusion Models with Deep Language UnderstandingChitwan Saharia, William Chan, Saurabh Saxena, Lala Li et al.NeurIPS 2022 · 8,965 citations
- Scalable Diffusion Models with TransformersWilliam Peebles, Saining XieICCV 2023 · 5,568 citations
Related papers
- RespoDiff: Dual-Module Bottleneck Transformation for Responsible & Faithful T2I GenerationSilpa Vadakkeeveetil Sreelatha, Sauradip Nag, Muhammad Awais, Serge J. Belongie et al.NeurIPS 2025 · 4 citations
- Self-Discovering Interpretable Diffusion Latent Directions for Responsible Text-to-Image GenerationHang Li, Chengzhi Shen, Philip Torr, Volker Tresp et al.CVPR 2024
- Responsible Text-to-Image Diffusion: Interpretable and Linearly Controllable Semantics for Fair and Safe GenerationSayedmoslem Shokrolahi, Jae-Mo Kang, Il-Min KimICML 2026
- Emergence and Evolution of Interpretable Concepts in Diffusion ModelsBerk Tinaz, Zalan Fabian, Mahdi SoltanolkotabiNeurIPS 2025 · 23 citations
- Training-Free Safe Text Embedding Guidance for Text-to-Image Diffusion ModelsByeonghu Na, Mina Kang, Jiseok Kwak, Minsang Park et al.NeurIPS 2025 · 8 citations
