Self-Supervised Sketch-to-Image Synthesis
Bingchen Liu, Yizhe Zhu, Kunpeng Song, Ahmed Elgammal
Abstract
Imagining a colored realistic image from an arbitrary drawn sketch is one of human capabilities that we eager machines to mimic. Unlike previous methods that either require the sketch-image pairs or utilize low-quantity detected edges as sketches, we study the exemplar-based sketch-to-image (s2i) synthesis task in a self-supervised learning manner, eliminating the necessity of the paired sketch data. To this end, we first propose an unsupervised method to efficiently synthesize line-sketches for general RGB-only datasets. With the synthetic paired-data, we then present a self-supervised Auto-Encoder (AE) to decouple the content/style features from sketches and RGB-images, and synthesize images that are both content-faithful to the sketches and style-consistent to the RGB-images. While prior works employ either the cycleconsistence loss or dedicated attentional modules to enforce the content/style fidelity, we show AE's superior performance with pure self-supervisions. To further improve the synthesis quality in high resolution, we also leverage an adversarial network to refine the details of the synthetic images. Extensive experiments on 1024 2 resolution demonstrate a new state-ofart-art performance of the proposed model on CelebA-HQ and Wiki-Art datasets. Moreover, with the proposed sketch generator, the model shows a promising performance on style mixing and style transfer, which require synthesized images to be both style-consistent and semantically meaningful. Our code is available on GitHub, and please visit Playform.io for an online demo of our model.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 78688c3c-5128-4d54-a032-8979f91a9dd4Cited by top-tier papers7
- High-Resolution Image Harmonization via Collaborative Dual TransformationsWenyan Cong, Xinhao Tao, Li Niu, Jing Liang et al.CVPR 2022 · 86 citations
- DeepFaceEditing: deep face generation and editing with disentangled geometry and appearance controlShu-Yu Chen, Feng-Lin Liu, Yu-Kun Lai, Paul L. Rosin et al.SIGGRAPH 2021 · 35 citations
- HAIGEN: Towards Human-AI Collaboration for Facilitating Creativity and Style Generation in Fashion DesignJianan Jiang, Di Wu, Hanhui Deng, Yidan Long et al.UbiComp 2024 · 28 citations
- Block and Detail: Scaffolding Sketch-to-Image GenerationVishnu Sarukkai, Lu Yuan, Mia Tang, Maneesh Agrawala et al.UIST 2024 · 23 citations
- Hierarchical Image Generation via Transformer-Based Sequential Patch SelectionXiaogang Xu, Ning XuAAAI 2022 · 10 citations
Builds on6
- U-GAT-IT: Unsupervised Generative Attentional Networks with Adaptive Layer-Instance Normalization for Image-to-Image TranslationJunho Kim, Minjae Kim, Hyeonwoo Kang, Kwanghee LeeICLR 2020 · 632 citations
- Cross-Domain Correspondence Learning for Exemplar-Based Image TranslationPan Zhang, Bo Zhang, Dong Chen, Lu Yuan et al.CVPR 2020
- MaskGAN: Towards Diverse and Interactive Facial Image ManipulationCheng-Han Lee, Ziwei Liu, Lingyun Wu, Ping LuoCVPR 2020
- Analyzing and Improving the Image Quality of StyleGANTero Karras, Samuli Laine, Miika Aittala, Janne Hellsten et al.CVPR 2020
- Reference-Based Sketch Image Colorization Using Augmented-Self Reference and Dense Semantic CorrespondenceJunsoo Lee, Eungyeup Kim, Yunsung Lee, Dongjun Kim et al.CVPR 2020
Related papers
- Stroke2Sketch: Harnessing Stroke Attributes for Training-Free Sketch GenerationRui Yang, Huining Li, Yiyi Long, Xiaojun Wu et al.ICCV 2025 · 2 citations
- Lab2Pix: Label-Adaptive Generative Adversarial Network for Unsupervised Image SynthesisLianli Gao, Junchen Zhu, Jingkuan Song, Feng Zheng et al.ACM MM 2020 · 12 citations
- Semi-supervised reference-based sketch extraction using a contrastive learning frameworkChang Wook Seo, Amirsaman Ashtari, Junyong NohSIGGRAPH 2023 · 14 citations
- Towards Controllable and Photorealistic Region-wise Image ManipulationAnsheng You, Chenglin Zhou, Qixuan Zhang, Lan XuACM MM 2021 · 2 citations
- SketchyCOCO: Image Generation From Freehand Scene SketchesChengying Gao, Qi Liu, Qi Xu, Limin Wang et al.CVPR 2020
