Zero-Shot Generative Model Adaptation via Image-Specific Prompt Learning
Jiayi Guo, Chaofei Wang, You Wu, Eric J. Zhang, Kai Wang, Xingqian Xu, Shiji Song, Humphrey Shi, Gao Huang
摘要
Ours "Wall painting" Ours "Anime painting" Ours "Photo" Source "Ukiyo-e" Ours NADA NADA NADA NADA Figure 1. The mode collapse issue. For NADA [21] and our method, the same generator pre-trained on the source domain of "Photo" is adapted to the unseen target domains of "Disney", "Anime painting", "Wall painting" and "Ukiyo-e" only with the domain labels. The images above the dotted line are some examples from the internet. The generated images of NADA exhibit some similar unseen patterns (yellow box areas) which are undesired in terms of quality and diversity. This issue is largely addressed by our method.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper13
- Rank-DETR for High Quality Object DetectionYifan Pu, Weicong Liang, Yiduo Hao, Yuhui Yuan 等NeurIPS 2023 · 被引用 138 次
- COVE: Unleashing the Diffusion Feature Correspondence for Consistent Video EditingJiangshan Wang, Yue Ma, Jiayi Guo, Yicheng Xiao 等NeurIPS 2024 · 被引用 76 次
- Prompt-Free Diffusion: Taking "Text" Out of Text-to-Image Diffusion ModelsXingqian Xu, Jiayi Guo, Zhangyang Wang, Gao Huang 等CVPR 2024 · 被引用 45 次
- Linear Differential Vision Transformer: Learning Visual Contrasts via Pairwise DifferentialsYifan Pu, Jixuan Ying, Qixiu Li, Tianzhu Ye 等NeurIPS 2025 · 被引用 9 次
- IMG: Calibrating Diffusion Models via Implicit Multimodal GuidanceJiayi Guo, Chuanhao Yan, Xingqian Xu, Yulin Wang 等ICCV 2025 · 被引用 4 次
它引用的顶会 Paper29
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- Denoising Diffusion Implicit ModelsJiaming Song, Chenlin Meng, Stefano ErmonICLR 2021 · 被引用 11,743 次
- Scaling Up Visual and Vision-Language Representation Learning With Noisy Text SupervisionChao Jia, Yinfei Yang, Ye Xia, Yi-Ting Chen 等ICML 2021 · 被引用 5,401 次
相关 Paper
- Few-Shot Image Generation via Cross-Domain CorrespondenceUtkarsh Ojha, Yijun Li, Jingwan Lu, Alexei A. Efros 等CVPR 2021
- Few-shot Cross-domain Image Generation via Inference-time Latent-code LearningArnab Kumar Mondal, Piyush Tiwary, Parag Singla, Prathosh APICLR 2023
- DATID-3D: Diversity-Preserved Domain Adaptation Using Text-to-Image Diffusion for 3D Generative ModelGwanghyun Kim, Se Young ChunCVPR 2023
- One-Shot Generative Domain AdaptationCeyuan Yang, Yujun Shen, Zhiyi Zhang, Yinghao Xu 等ICCV 2023 · 被引用 69 次
- UniGAN: Reducing Mode Collapse in GANs using a Uniform GeneratorZiqi Pan, Li Niu, Liqing ZhangNeurIPS 2022 · 被引用 17 次
