Named Entity Driven Zero-Shot Image Manipulation
Zhida Feng, Li Chen, Jing Tian, Jiaxiang Liu, Shikun Feng
2024Year
Abstract
looks like Joe Biden 10-years-old female wearing lipstick with curly hair with goatee Figure 1. Zero-Shot Manipulation with StyleEntity. Top row: Input images. Bottom row: Manipulation results. All prompts are unseen in the training set.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Builds on33
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 24,064 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- Photorealistic Text-to-Image Diffusion Models with Deep Language UnderstandingChitwan Saharia, William Chan, Saurabh Saxena, Lala Li et al.NeurIPS 2022 · 8,965 citations
- GLIDE: Towards Photorealistic Image Generation and Editing with Text-Guided Diffusion ModelsAlexander Quinn Nichol, Prafulla Dhariwal, Aditya Ramesh, Pranav Shyam et al.ICML 2022 · 4,691 citations
Related papers
- When StyleGAN Meets Stable Diffusion: a Adapter for Personalized Image GenerationXiaoming Li, Xinyu Hou, Chen Change LoyCVPR 2024
- Playmate: Flexible Control of Portrait Animation via 3D-Implicit Space Guided DiffusionXingpei Ma, Jiaran Cai, Yuansheng Guan, Shenneng Huang et al.ICML 2025
- L2M-GAN: Learning To Manipulate Latent Space Semantics for Facial Attribute EditingGuoxing Yang, Nanyi Fei, Mingyu Ding, Guangzhen Liu et al.CVPR 2021
- A Latent Transformer for Disentangled Face Editing in Images and VideosXu Yao, Alasdair Newson, Yann Gousseau, Pierre HellierICCV 2021 · 97 citations
- TextCraftor: Your Text Encoder can be Image Quality ControllerYanyu Li, Xian Liu, Anil Kag, Ju Hu et al.CVPR 2024
