Improving Subject-Driven Image Synthesis with Subject-Agnostic Guidance
Kelvin C. K. Chan, Yang Zhao, Xuhui Jia, Ming-Hsuan Yang, Huisheng Wang
2024年份
2顶会引用
摘要
style w/ SAG S*, swimming in front of Eiffel Tower, Van Gogh starry night style w/o SAG S* in a basket, at a beach S* with a cloudy night sky and a moon S*, Pixar movie Reference Figure 1. Addressing Content Ignorance. Given user-provided subject images, a part of the content specified in the text prompt (highlighted in blue) are overlooked. Our Subject-Agnostic Guidance (SAG) aligns the output more closely with both the target subject and text prompt. Here S * denotes a pseudo-word, with its text embedding replaced by a learnable subject embedding.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- Steering Guidance for Personalized Text-to-Image Diffusion ModelsSunghyun Park, Seokeon Choi, Hyoungwoo Park, Sungrack YunICCV 2025 · 被引用 2 次
- PosterMaker: Towards High-Quality Product Poster Generation with Accurate Text RenderingYifan Gao, Zihang Lin, Chuanbin Liu, Min Zhou 等CVPR 2025
它引用的顶会 Paper28
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- Diffusion Models Beat GANs on Image SynthesisPrafulla Dhariwal, Alexander Quinn NicholNeurIPS 2021 · 被引用 13,211 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- Photorealistic Text-to-Image Diffusion Models with Deep Language UnderstandingChitwan Saharia, William Chan, Saurabh Saxena, Lala Li 等NeurIPS 2022 · 被引用 8,965 次
相关 Paper
- ArtAdapter: Text-to-Image Style Transfer using Multi-Level Style Encoder and Explicit AdaptationDar-Yen Chen, Hamish Tennent, Ching-Wen HsuCVPR 2024
- LAPIG: Language Guided Projector Image Generation with Surface Adaptation and StylizationYuchen Deng, Haibin Ling, Bingyao HuangIEEE VR 2025 · 被引用 6 次
- Event-Customized Image GenerationZhen Wang, Yilei Jiang, Dong Zheng, Jun Xiao 等ICML 2025
- StyleStudio: Text-Driven Style Transfer with Selective Control of Style ElementsMingkun Lei, Xue Song, Beier Zhu, Hao Wang 等CVPR 2025
- A Training-Free Style-Personalization via SVD-Based Feature DecompositionKyoungmin Lee, Jihun Park, Jongmin Gim, Wonhyeok Choi 等CVPR 2026
