Tag2Pix: Line Art Colorization Using Text Tag With SECat and Changing Loss
Hyunsu Kim, Ho Young Jhoo, Eunhyeok Park, Sungjoo Yoo
摘要
Line art colorization is expensive and challenging to automate. A GAN approach is proposed, called Tag2Pix, of line art colorization which takes as input a grayscale line art and color tag information and produces a quality colored image. First, we present the Tag2Pix line art colorization dataset. A generator network is proposed which consists of convolutional layers to transform the input line art, a pre-trained semantic extraction network, and an encoder for input color information. The discriminator is based on an auxiliary classifier GAN to classify the tag information as well as genuineness. In addition, we propose a novel network structure called SECat, which makes the generator properly colorize even small features such as eyes, and also suggest a novel two-step training method where the generator and discriminator first learn the notion of object and shape and then, based on the learned notion, learn colorization, such as where and how to place which color. We present both quantitative and qualitative evaluations which prove the effectiveness of the proposed method.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper23
- Frequency Domain Image Translation: More Photo-realistic, Better Identity-preservingMu Cai, Hong Zhang, Huijuan Huang, Qichuan Geng 等ICCV 2021 · 被引用 118 次
- LapsCore: Language-guided Person Search via Color ReasoningYushuang Wu, Zizheng Yan, Xiaoguang Han, Guanbin Li 等ICCV 2021 · 被引用 91 次
- Feature Statistics Mixing Regularization for Generative Adversarial NetworksJunho Kim, Yunjey Choi, Youngjung UhCVPR 2022 · 被引用 22 次
- Semi-supervised reference-based sketch extraction using a contrastive learning frameworkChang Wook Seo, Amirsaman Ashtari, Junyong NohSIGGRAPH 2023 · 被引用 14 次
- Style-Structure Disentangled Features and Normalizing Flows for Diverse Icon ColorizationYuan-kui Li, Yun-Hsuan Lien, Yu-Shuen WangCVPR 2022 · 被引用 13 次
相关 Paper
- DACoN: DINO for Anime Paint Bucket Colorization with Any Number of Reference ImagesKazuma Nagata, Naoshi KanekoICCV 2025 · 被引用 1 次
- Repurposing GANs for One-Shot Semantic Part SegmentationNontawat Tritrong, Pitchaporn Rewatbowornwong, Supasorn SuwajanakornCVPR 2021
- LineArt: A Knowledge-guided Training-free High-quality Appearance Transfer for Design Drawing with Diffusion ModelXi Wang, Hongzhen Li, Heng Fang, Yichen Peng 等CVPR 2025
- FlatGAN: A Holistic Approach for Robust Flat-Coloring in High-Definition with Understanding Line DiscontinuityHan Kim, Chunggi Lee, Junsoo Lee, Dohyun Kim 等ACM MM 2023 · 被引用 1 次
- Lab2Pix: Label-Adaptive Generative Adversarial Network for Unsupervised Image SynthesisLianli Gao, Junchen Zhu, Jingkuan Song, Feng Zheng 等ACM MM 2020 · 被引用 12 次
