SketchDeco: Training-Free Latent Composition for Precise Sketch Colourisation
Chaitat Utintu, Yi-Zhe Song
摘要
We introduce SketchDeco, a training-free approach to sketch colourisation that bridges the gap between professional design needs and intuitive, region-based control. Our method empowers artists to use simple masks and colour palettes for precise spatial and chromatic specification, avoiding both the tediousness of manual assignment and the ambiguity of text-based prompts. We reformulate this task as a novel, training-free composition problem. Our core technical contribution is a guided latent-space blending process: we first leverage diffusion inversion to precisely ``paint'' user-defined colours into specified regions, and then use a custom self-attention mechanism to harmoniously blend these local edits with a globally consistent base image. This ensures both local colour fidelity and global harmony without requiring any model fine-tuning. Our system produces high-quality results in 15--20 inference steps on consumer GPUs, making professional-quality, controllable colourisation accessible.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- FusAIn: Composing Generative AI Visual Prompts Using Pen-based InteractionXiaohan Peng, Janin Koch, Wendy E. MackayCHI 2025 · 被引用 15 次
- Magiccolor: Multi-Instance Sketch ColorizationYinhan Zhang, Yue Ma, Bingyuan Wang, Qifeng Chen 等ICCV 2025 · 被引用 3 次
- VQ-SGen: A Vector Quantized Stroke Representation for Creative Sketch GenerationJiawei Wang, Zhiming Cui, Changjian LiICCV 2025 · 被引用 3 次
- Cobra: Efficient Line Art COlorization with BRoAder ReferencesJunhao Zhuang, Lingen Li, Xuan Ju, Zhaoyang Zhang 等SIGGRAPH 2025 · 被引用 2 次
它引用的顶会 Paper39
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- Diffusion Models Beat GANs on Image SynthesisPrafulla Dhariwal, Alexander Quinn NicholNeurIPS 2021 · 被引用 13,211 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- Denoising Diffusion Implicit ModelsJiaming Song, Chenlin Meng, Stefano ErmonICLR 2021 · 被引用 11,743 次
相关 Paper
- Text to Sketch Generation with Multi-StylesTengjie Li, Shikui Tu, Lei XuNeurIPS 2025 · 被引用 1 次
- FreeControl: Efficient, Training-Free Structural Control via One-Step Attention ExtractionJiang Lin, Xinyu Chen, Song Wu, Zhiqiu Zhang 等NeurIPS 2025 · 被引用 3 次
- Stroke2Sketch: Harnessing Stroke Attributes for Training-Free Sketch GenerationRui Yang, Huining Li, Yiyi Long, Xiaojun Wu 等ICCV 2025 · 被引用 2 次
- Versatile Vision Foundation Model for Image and Video ColorizationVukasin Bozic, Abdelaziz Djelouah, Yang Zhang, Radu Timofte 等SIGGRAPH 2024 · 被引用 9 次
- Draw2Edit: Mask-Free Sketch-Guided Image ManipulationYiwen Xu, Ruoyu Guo, Maurice Pagnucco, Yang SongACM MM 2023 · 被引用 3 次
