Tag2Pix: Line Art Colorization Using Text Tag With SECat and Changing Loss
Hyunsu Kim, Ho Young Jhoo, Eunhyeok Park, Sungjoo Yoo
Abstract
Line art colorization is expensive and challenging to automate. A GAN approach is proposed, called Tag2Pix, of line art colorization which takes as input a grayscale line art and color tag information and produces a quality colored image. First, we present the Tag2Pix line art colorization dataset. A generator network is proposed which consists of convolutional layers to transform the input line art, a pre-trained semantic extraction network, and an encoder for input color information. The discriminator is based on an auxiliary classifier GAN to classify the tag information as well as genuineness. In addition, we propose a novel network structure called SECat, which makes the generator properly colorize even small features such as eyes, and also suggest a novel two-step training method where the generator and discriminator first learn the notion of object and shape and then, based on the learned notion, learn colorization, such as where and how to place which color. We present both quantitative and qualitative evaluations which prove the effectiveness of the proposed method.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 8a0c9e94-7dc2-450c-927e-8378251bfe95Cited by top-tier papers23
- Frequency Domain Image Translation: More Photo-realistic, Better Identity-preservingMu Cai, Hong Zhang, Huijuan Huang, Qichuan Geng et al.ICCV 2021 · 118 citations
- LapsCore: Language-guided Person Search via Color ReasoningYushuang Wu, Zizheng Yan, Xiaoguang Han, Guanbin Li et al.ICCV 2021 · 91 citations
- Feature Statistics Mixing Regularization for Generative Adversarial NetworksJunho Kim, Yunjey Choi, Youngjung UhCVPR 2022 · 22 citations
- Semi-supervised reference-based sketch extraction using a contrastive learning frameworkChang Wook Seo, Amirsaman Ashtari, Junyong NohSIGGRAPH 2023 · 14 citations
- Style-Structure Disentangled Features and Normalizing Flows for Diverse Icon ColorizationYuan-kui Li, Yun-Hsuan Lien, Yu-Shuen WangCVPR 2022 · 13 citations
Related papers
- DACoN: DINO for Anime Paint Bucket Colorization with Any Number of Reference ImagesKazuma Nagata, Naoshi KanekoICCV 2025 · 1 citation
- Repurposing GANs for One-Shot Semantic Part SegmentationNontawat Tritrong, Pitchaporn Rewatbowornwong, Supasorn SuwajanakornCVPR 2021
- LineArt: A Knowledge-guided Training-free High-quality Appearance Transfer for Design Drawing with Diffusion ModelXi Wang, Hongzhen Li, Heng Fang, Yichen Peng et al.CVPR 2025
- FlatGAN: A Holistic Approach for Robust Flat-Coloring in High-Definition with Understanding Line DiscontinuityHan Kim, Chunggi Lee, Junsoo Lee, Dohyun Kim et al.ACM MM 2023 · 1 citation
- Lab2Pix: Label-Adaptive Generative Adversarial Network for Unsupervised Image SynthesisLianli Gao, Junchen Zhu, Jingkuan Song, Feng Zheng et al.ACM MM 2020 · 12 citations
