Learning to Incorporate Texture Saliency Adaptive Attention to Image Cartoonization
Xiang Gao, Yuqi Zhang, Yingjie Tian
Abstract
Image cartoonization is recently dominated by generative adversarial networks (GANs) from the perspective of unsupervised image-to-image translation, in which an inherent challenge is to precisely capture and sufficiently transfer characteristic cartoon styles (e.g., clear edges, smooth color shading, abstract fine structures, etc.). Existing advanced models try to enhance cartoonization effect by learning to promote edges adversarially, introducing style transfer loss, or learning to align style from multiple representation space. This paper demonstrates that more distinct and vivid cartoonization effect could be easily achieved with only basic adversarial loss. Observing that cartoon style is more evident in cartoon-texture-salient local image regions, we build a region-level adversarial learning branch in parallel with the normal image-level one, which constrains adversarial learning on cartoon-texture-salient local patches for better perceiving and transferring cartoon texture features. To this end, a novel cartoon-texture-saliency-sampler (CTSS) module is proposed to dynamically sample cartoon-texture-salient patches from training data. With extensive experiments, we demonstrate that texture saliency adaptive attention in adversarial learning, as a missing ingredient of related methods in image cartoonization, is of significant importance in facilitating and enhancing image cartoon stylization, especially for high-resolution input pictures. Our code is publically available at this github link.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext db8fce1c-025b-4324-9e84-e3304d08f535Cited by top-tier papers3
- ArtBank: Artistic Style Transfer with Pre-trained Diffusion Model and Implicit Style Prompt BankZhanjie Zhang, Quanwei Zhang, Wei Xing, Guangyuan Li et al.AAAI 2024 · 32 citations
- Scenimefy: Learning to Craft Anime Scene via Semi-Supervised Image-to-Image TranslationYuxin Jiang, Liming Jiang, Shuai Yang, Chen Change LoyICCV 2023 · 25 citations
- Frequency-Controlled Diffusion Model for Versatile Text-Guided Image-to-Image TranslationXiang Gao, Zhengbo Xu, Junhan Zhao, Jiaying LiuAAAI 2024 · 23 citations
Builds on2
Related papers
- Cartoon-Flow: A Flow-Based Generative Adversarial Network for Arbitrary-Style Photo CartoonizationJieun Lee, Hyeonwoo Kim, Jonghwa Shim, Eenjun HwangACM MM 2022 · 14 citations
- TSSAT: Two-Stage Statistics-Aware Transformation for Artistic Style TransferHaibo Chen, Lei Zhao, Jun Li, Jian YangACM MM 2023 · 21 citations
- Interactive Cartoonization with Controllable Perceptual FactorsNamhyuk Ahn, Patrick Kwon, Jihye Back, Kibeom Hong et al.CVPR 2023
- Unsupervised Coherent Video Cartoonization with Perceptual Motion ConsistencyZhenhuan Liu, Liang Li, Huajie Jiang, Xin Jin et al.AAAI 2022 · 7 citations
- SCSA: A Plug-and-Play Semantic Continuous-Sparse Attention for Arbitrary Semantic Style TransferChunnan Shang, Zhizhong Wang, Hongwei Wang, Xiangming MengCVPR 2025
