Visual Instruction Inversion: Image Editing via Image Prompting
Thao Nguyen, Yuheng Li, Utkarsh Ojha, Yong Jae Lee
2023Year
53Citations
19Top-tier citations
Abstract
Test Image Before After "Turn it into a watercolor painting" Instruction + "… husky" + "… pit bull" + "… tiger" + "… monkey" + "… squirrel" Instruction InstructPix2Pix Ours Test Image <ins> + "… husky" + "… pit bull" + "… tiger"
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers19
- ImageRAG: Dynamic Image Retrieval for Reference-Guided Image GenerationRotem Shalev-Arkushin, Rinon Gal, Amit Bermano, Ohad FriedICLR 2026 · 25 citations
- ParallelEdits: Efficient Multi-Aspect Text-Driven Image Editing with Attention GroupingMingzhen Huang, Jialing Cai, Shan Jia, Vishnu Suresh Lokhande et al.NeurIPS 2024 · 19 citations
- OneActor: Consistent Subject Generation via Cluster-Conditioned GuidanceJiahao Wang, Caixia Yan, Haonan Lin, Weizhan Zhang et al.NeurIPS 2024 · 16 citations
- 3DOT: Texture Transfer for 3DGS Objects from a Single Reference ImageXiao Cao, Beibei Lin, Bo Wang, Zhiyong Huang et al.NeurIPS 2025 · 8 citations
- Analogist: Out-of-the-box Visual In-Context Learning with Image Diffusion ModelZheng Gu, Shiyuan Yang, Jing Liao, Jing Huo et al.SIGGRAPH 2024 · 7 citations
Builds on33
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah et al.NeurIPS 2020 · 64,255 citations
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- Diffusion Models Beat GANs on Image SynthesisPrafulla Dhariwal, Alexander Quinn NicholNeurIPS 2021 · 13,211 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
Related papers
- UniGen-1.5: Enhancing Image Generation and Editing through Reward Unification in RLRui Tian, Mingfei Gao, Haiming Gang, Jiasen Lu et al.CVPR 2026
- Multi-Concept Customization of Text-to-Image DiffusionNupur Kumari, Bingliang Zhang, Richard Zhang, Eli Shechtman et al.CVPR 2023
- SmartEdit: Exploring Complex Instruction-Based Image Editing with Multimodal Large Language ModelsYuzhou Huang, Liangbin Xie, Xintao Wang, Ziyang Yuan et al.CVPR 2024
- HIVE: Harnessing Human Feedback for Instructional Visual EditingShu Zhang, Xinyi Yang, Yihao Feng, Can Qin et al.CVPR 2024
- IMAGINE: Image Synthesis by Image-Guided Model InversionPei Wang, Yijun Li, Krishna Kumar Singh, Jingwan Lu et al.CVPR 2021
