Named Entity Driven Zero-Shot Image Manipulation
Zhida Feng, Li Chen, Jing Tian, Jiaxiang Liu, Shikun Feng
2024年份
摘要
looks like Joe Biden 10-years-old female wearing lipstick with curly hair with goatee Figure 1. Zero-Shot Manipulation with StyleEntity. Top row: Input images. Bottom row: Manipulation results. All prompts are unseen in the training set.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper33
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 被引用 24,064 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- Photorealistic Text-to-Image Diffusion Models with Deep Language UnderstandingChitwan Saharia, William Chan, Saurabh Saxena, Lala Li 等NeurIPS 2022 · 被引用 8,965 次
- GLIDE: Towards Photorealistic Image Generation and Editing with Text-Guided Diffusion ModelsAlexander Quinn Nichol, Prafulla Dhariwal, Aditya Ramesh, Pranav Shyam 等ICML 2022 · 被引用 4,691 次
相关 Paper
- When StyleGAN Meets Stable Diffusion: a Adapter for Personalized Image GenerationXiaoming Li, Xinyu Hou, Chen Change LoyCVPR 2024
- Playmate: Flexible Control of Portrait Animation via 3D-Implicit Space Guided DiffusionXingpei Ma, Jiaran Cai, Yuansheng Guan, Shenneng Huang 等ICML 2025
- L2M-GAN: Learning To Manipulate Latent Space Semantics for Facial Attribute EditingGuoxing Yang, Nanyi Fei, Mingyu Ding, Guangzhen Liu 等CVPR 2021
- A Latent Transformer for Disentangled Face Editing in Images and VideosXu Yao, Alasdair Newson, Yann Gousseau, Pierre HellierICCV 2021 · 被引用 97 次
- TextCraftor: Your Text Encoder can be Image Quality ControllerYanyu Li, Xian Liu, Anil Kag, Ju Hu 等CVPR 2024
