Lune

CVPR2026Top-tier venue

MagicQuill V2: Precise and Interactive Image Editing with Layered Visual Cues

Zichen Liu, Yue Yu, Hao Ouyang, Qiuyu Wang, Shuailei Ma, Ka Leong Cheng, Wen Wang, Qingyan Bai, Yuxuan Zhang, Yanhong Zeng, Yixuan Li, Xing Zhu

2026Year

Abstract

Figure 1. MagicQuill V2 introduces a layered composition framework for precise generative image editing. Users articulate complex intents by stacking independent visual layers in a continuous workflow: (A) composing a base scene with multiple content layers (car, lady, dog); (B) restructuring the dog's pose with a spatial layer; (C) modifying visual attributes by sketching a bell collar on the structural layer and painting the shirt on the color layer; (D) inserting final details like a hat and apple seamlessly. Demo available at project page.

Ask about this paper

Your agent reads all of it.

Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.

Questions to start from

Your agent calls

Luneget_paper_fulltext

Ask in Lune

Free to start. No credit card required.

lune papers fulltext 3d95375b-2107-4e45-9cbf-8da75b1409b7

Builds on36

Related papers

Dusk over the sea between two cliffs drawn in fine vertical lines