Text-Driven Image Editing via Learnable Regions
Yuanze Lin, Yi-Wen Chen, Yi-Hsuan Tsai, Lu Jiang, Ming-Hsuan Yang
2024Year
13Top-tier citations
Abstract
donuts with ice cream a cute sloth holds a box a flowering cherry tree a blooming flower and dessert several lotus flowers growing in the water a cup of coffee next to the bread Figure 1. Overview. Given an input image and a language description for editing, our method can generate realistic and relevant images without the need for user-specified regions for editing. It performs local image editing while preserving the image context.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 5918689c-bf43-44ca-8ddf-b3be742d876eCited by top-tier papers13
- DiffVax: Optimization-Free Image Immunization Against Diffusion-Based EditingTarik Can Ozden, Ozgur Kara, Oguzhan Akcin, Kerem Zaman et al.ICLR 2026 · 7 citations
- IllumiCraft: Unified Geometry and Illumination Diffusion for Controllable Video GenerationYuanze Lin, Yi-Wen Chen, Yi-Hsuan Tsai, Ronald Clark et al.NeurIPS 2025 · 6 citations
- Textualize Visual Prompt for Image Editing via Diffusion BridgePengcheng Xu, Qingnan Fan, Fei Kou, Shuai Qin et al.AAAI 2025 · 4 citations
- PartEdit: Fine-Grained Image Editing using Pre-Trained Diffusion ModelsAleksandar Cvejic, Abdelrahman Eldesokey, Peter WonkaSIGGRAPH 2025 · 3 citations
- SuperEdit: Rectifying and Facilitating Supervision for Instruction-Based Image EditingMing Li, Xin Gu, Fan Chen, Xiaoying Xing et al.ICCV 2025 · 2 citations
Builds on39
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah et al.NeurIPS 2020 · 64,255 citations
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- Diffusion Models Beat GANs on Image SynthesisPrafulla Dhariwal, Alexander Quinn NicholNeurIPS 2021 · 13,211 citations
Related papers
- InstructPix2Pix: Learning to Follow Image Editing InstructionsTim Brooks, Aleksander Holynski, Alexei A. EfrosCVPR 2023
- Blended Diffusion for Text-driven Editing of Natural ImagesOmri Avrahami, Dani Lischinski, Ohad FriedCVPR 2022 · 670 citations
- IntrinsicEdit: Precise generative image manipulation in intrinsic spaceLinjie Lyu, Valentin Deschaintre, Yannick Hold-Geoffroy, Milos Hasan et al.SIGGRAPH 2025 · 7 citations
- Fully Functional Image Manipulation Using Scene Graphs in A Bounding-Box Free WaySitong Su, Lianli Gao, Junchen Zhu, Jie Shao et al.ACM MM 2021 · 16 citations
- OBJECT 3DIT: Language-guided 3D-aware Image EditingOscar Michel, Anand Bhattad, Eli VanderBilt, Ranjay Krishna et al.NeurIPS 2023 · 79 citations
