Zero-Shot Generative Model Adaptation via Image-Specific Prompt Learning
Jiayi Guo, Chaofei Wang, You Wu, Eric J. Zhang, Kai Wang, Xingqian Xu, Shiji Song, Humphrey Shi, Gao Huang
Abstract
Ours "Wall painting" Ours "Anime painting" Ours "Photo" Source "Ukiyo-e" Ours NADA NADA NADA NADA Figure 1. The mode collapse issue. For NADA [21] and our method, the same generator pre-trained on the source domain of "Photo" is adapted to the unseen target domains of "Disney", "Anime painting", "Wall painting" and "Ukiyo-e" only with the domain labels. The images above the dotted line are some examples from the internet. The generated images of NADA exhibit some similar unseen patterns (yellow box areas) which are undesired in terms of quality and diversity. This issue is largely addressed by our method.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext c154743d-4904-4836-96f7-4b04ab02cd96Cited by top-tier papers13
- Rank-DETR for High Quality Object DetectionYifan Pu, Weicong Liang, Yiduo Hao, Yuhui Yuan et al.NeurIPS 2023 · 138 citations
- COVE: Unleashing the Diffusion Feature Correspondence for Consistent Video EditingJiangshan Wang, Yue Ma, Jiayi Guo, Yicheng Xiao et al.NeurIPS 2024 · 76 citations
- Prompt-Free Diffusion: Taking "Text" Out of Text-to-Image Diffusion ModelsXingqian Xu, Jiayi Guo, Zhangyang Wang, Gao Huang et al.CVPR 2024 · 45 citations
- Linear Differential Vision Transformer: Learning Visual Contrasts via Pairwise DifferentialsYifan Pu, Jixuan Ying, Qixiu Li, Tianzhu Ye et al.NeurIPS 2025 · 9 citations
- IMG: Calibrating Diffusion Models via Implicit Multimodal GuidanceJiayi Guo, Chuanhao Yan, Xingqian Xu, Yulin Wang et al.ICCV 2025 · 4 citations
Builds on29
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- Denoising Diffusion Implicit ModelsJiaming Song, Chenlin Meng, Stefano ErmonICLR 2021 · 11,743 citations
- Scaling Up Visual and Vision-Language Representation Learning With Noisy Text SupervisionChao Jia, Yinfei Yang, Ye Xia, Yi-Ting Chen et al.ICML 2021 · 5,401 citations
Related papers
- Few-Shot Image Generation via Cross-Domain CorrespondenceUtkarsh Ojha, Yijun Li, Jingwan Lu, Alexei A. Efros et al.CVPR 2021
- Few-shot Cross-domain Image Generation via Inference-time Latent-code LearningArnab Kumar Mondal, Piyush Tiwary, Parag Singla, Prathosh APICLR 2023
- DATID-3D: Diversity-Preserved Domain Adaptation Using Text-to-Image Diffusion for 3D Generative ModelGwanghyun Kim, Se Young ChunCVPR 2023
- One-Shot Generative Domain AdaptationCeyuan Yang, Yujun Shen, Zhiyi Zhang, Yinghao Xu et al.ICCV 2023 · 69 citations
- UniGAN: Reducing Mode Collapse in GANs using a Uniform GeneratorZiqi Pan, Li Niu, Liqing ZhangNeurIPS 2022 · 17 citations
