Improving Subject-Driven Image Synthesis with Subject-Agnostic Guidance
Kelvin C. K. Chan, Yang Zhao, Xuhui Jia, Ming-Hsuan Yang, Huisheng Wang
Abstract
style w/ SAG S*, swimming in front of Eiffel Tower, Van Gogh starry night style w/o SAG S* in a basket, at a beach S* with a cloudy night sky and a moon S*, Pixar movie Reference Figure 1. Addressing Content Ignorance. Given user-provided subject images, a part of the content specified in the text prompt (highlighted in blue) are overlooked. Our Subject-Agnostic Guidance (SAG) aligns the output more closely with both the target subject and text prompt. Here S * denotes a pseudo-word, with its text embedding replaced by a learnable subject embedding.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext e4489b3f-abb4-4b75-aeaf-8d420a21bcd0Cited by top-tier papers2
- Steering Guidance for Personalized Text-to-Image Diffusion ModelsSunghyun Park, Seokeon Choi, Hyoungwoo Park, Sungrack YunICCV 2025 · 2 citations
- PosterMaker: Towards High-Quality Product Poster Generation with Accurate Text RenderingYifan Gao, Zihang Lin, Chuanbin Liu, Min Zhou et al.CVPR 2025
Builds on28
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- Diffusion Models Beat GANs on Image SynthesisPrafulla Dhariwal, Alexander Quinn NicholNeurIPS 2021 · 13,211 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- Photorealistic Text-to-Image Diffusion Models with Deep Language UnderstandingChitwan Saharia, William Chan, Saurabh Saxena, Lala Li et al.NeurIPS 2022 · 8,965 citations
Related papers
- ArtAdapter: Text-to-Image Style Transfer using Multi-Level Style Encoder and Explicit AdaptationDar-Yen Chen, Hamish Tennent, Ching-Wen HsuCVPR 2024
- LAPIG: Language Guided Projector Image Generation with Surface Adaptation and StylizationYuchen Deng, Haibin Ling, Bingyao HuangIEEE VR 2025 · 6 citations
- Event-Customized Image GenerationZhen Wang, Yilei Jiang, Dong Zheng, Jun Xiao et al.ICML 2025
- StyleStudio: Text-Driven Style Transfer with Selective Control of Style ElementsMingkun Lei, Xue Song, Beier Zhu, Hao Wang et al.CVPR 2025
- A Training-Free Style-Personalization via SVD-Based Feature DecompositionKyoungmin Lee, Jihun Park, Jongmin Gim, Wonhyeok Choi et al.CVPR 2026
