Improving Editability in Image Generation with Layer-wise Memory
Daneul Kim, Jaeah Lee, Jaesik Park
2025Year
1Top-tier citations
Abstract
HD Printer BLD Photoshop 2025 Pincel Image Inpainting Models Commercial Products "A forest" + "Lego man standing" + "Jeep, front view" + "A dog sitting" Mask 3 Mask 1 Ours O r d e r Figure 1. Overview. Our framework enables the interactive generation of images with enhanced control but in a simple manner, by rough mask and prompt, through iterative scene editing. We utilize the background scene generated by our framework to edit in HD Painter [32] or Blended Latent Diffusion (BLD) [3] for comparison and commercial products like Photoshop [1] and Pincel [37].
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers1
Ask how each one uses itBuilds on34
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- Visual Instruction TuningHaotian Liu, Chunyuan Li, Qingyang Wu, Yong Jae LeeNeurIPS 2023 · 11,349 citations
- Adding Conditional Control to Text-to-Image Diffusion ModelsLvmin Zhang, Anyi Rao, Maneesh AgrawalaICCV 2023 · 6,759 citations
- GLIDE: Towards Photorealistic Image Generation and Editing with Text-Guided Diffusion ModelsAlexander Quinn Nichol, Prafulla Dhariwal, Aditya Ramesh, Pranav Shyam et al.ICML 2022 · 4,691 citations
Related papers
- Streamlining Image Editing with Layered Diffusion BrushesPeyman Gholami, Robert XiaoICCV 2025 · 1 citation
- IntrinsicEdit: Precise generative image manipulation in intrinsic spaceLinjie Lyu, Valentin Deschaintre, Yannick Hold-Geoffroy, Milos Hasan et al.SIGGRAPH 2025 · 7 citations
- MirrorVerse: Pushing Diffusion Models to Realistically Reflect the WorldAnkit Dhiman, Manan Shah, R. Venkatesh BabuCVPR 2025
- Spin: Diffusion-based Semantic Image Painting Through Independent Information InjectionDantong Wu, Zhiqiang Chen, Tianjiao Du, Peipei Ran et al.AAAI 2025
- VISION-XL: High Definition Video Inverse Problem Solver using Latent Image Diffusion ModelsTaesung Kwon, Jong Chul YeICCV 2025
