MagicQuill: An Intelligent Interactive Image Editing System
Zichen Liu, Yue Yu, Hao Ouyang, Qiuyu Wang, Ka Leong Cheng, Wen Wang, Zhiheng Liu, Qifeng Chen, Yujun Shen
Abstract
Ant Group, 3 ZJU, 4 HKU Figure 1 . MagicQuill is an intelligent and interactive image editing system built upon diffusion models. Users seamlessly edit images using three intuitive brushstrokes: add, subtract, and color (A). A MLLM dynamically predicts user intentions from their brush strokes and suggests contextual prompts (B1-B4). The examples demonstrate diverse editing operations: to generate a jacket from clothing contour (B1), add a flower crown from head sketches (B2), remove background (B3), and apply color changes to the hair and flowers (B4).
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext de473f42-149a-49c5-b5b1-40d7fc836850Cited by top-tier papers11
- ContextFlow: Training-Free Video Object Editing via Adaptive Context EnrichmentYiyang Chen, Xuanhua He, Xiujun Ma, Jack MaAAAI 2026 · 17 citations
- DesignTrace: Exploring, Iterating and Tracking Design Alternatives with GenAIXiaohan Peng, Debanjana Haldar, Wendy E. Mackay, Janin KochCHI 2026 · 3 citations
- Interaction-Augmented Instruction: Modeling the Synergy of Prompts and Interactions in Human-GenAI CollaborationLeixian Shen, Yifang Wang, Huamin Qu, Xing Xie et al.CHI 2026 · 3 citations
- Sketch3DVE: Sketch-based 3D-Aware Scene Video EditingFeng-Lin Liu, Shi-Yang Li, Yan-Pei Cao, Hongbo Fu et al.SIGGRAPH 2025 · 2 citations
- Live Interactive Training for Video SegmentationXinyu Yang, Haozheng Yu, Yihong Sun, Bharath Hariharan et al.CVPR 2026 · 1 citation
Builds on39
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- Denoising Diffusion Implicit ModelsJiaming Song, Chenlin Meng, Stefano ErmonICLR 2021 · 11,743 citations
- Visual Instruction TuningHaotian Liu, Chunyuan Li, Qingyang Wu, Yong Jae LeeNeurIPS 2023 · 11,349 citations
Related papers
- MagicQuill V2: Precise and Interactive Image Editing with Layered Visual CuesZichen Liu, Yue Yu, Hao Ouyang, Qiuyu Wang et al.CVPR 2026
- ImageBrush: Learning Visual In-Context Instructions for Exemplar-Based Image ManipulationYasheng Sun, Yifan Yang, Houwen Peng, Yifei Shen et al.NeurIPS 2023 · 71 citations
- MagicPaint: Operate Anything for Image Inpainting with Diffusion ModelQinhong Yang, Dongdong Chen, Qi Chu, Tao Gong et al.AAAI 2026
- Zero-shot Image Editing with Reference ImitationXi Chen, Yutong Feng, Mengting Chen, Yiyang Wang et al.NeurIPS 2024 · 80 citations
- An Item Is Worth a Prompt: Versatile Image Editing with Disentangled ControlAosong Feng, Weikang Qiu, Jinbin Bai, Zhen Dong et al.AAAI 2025 · 9 citations
