Scenimefy: Learning to Craft Anime Scene via Semi-Supervised Image-to-Image Translation
Yuxin Jiang, Liming Jiang, Shuai Yang, Chen Change Loy
摘要
Automatic high-quality rendering of anime scenes from complex real-world images is of significant practical value. The challenges of this task lie in the complexity of the scenes, the unique features of anime style, and the lack of high-quality datasets to bridge the domain gap. Despite promising attempts, previous efforts are still incompetent in achieving satisfactory results with consistent semantic preservation, evident stylization, and fine details. In this study, we propose Scenimefy, a novel semisupervised image-to-image translation framework that addresses these challenges. Our approach guides the learning with structure-consistent pseudo paired data, simplifying the pure unsupervised setting. The pseudo data are derived uniquely from a semantic-constrained StyleGAN leveraging rich model priors like CLIP. We further apply segmentationguided data selection to obtain high-quality pseudo supervision. A patch-wise contrastive style loss is introduced to improve stylization and fine details. Besides, we contribute a high-resolution anime scene dataset to facilitate future research. Our extensive experiments demonstrate the su- * Equal contribution. periority of our method over state-of-the-art baselines in terms of both perceptual quality and quantitative performance. Project page: https://yuxinn-j.github. io/projects/Scenimefy.html .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- ArtEditor: Learning Customized Instructional Image Editor From Few-Shot ExamplesShijie Huang, Yiren Song, Yuxuan Zhang, Hailong Guo 等ICCV 2025 · 被引用 2 次
- Balanced Image Stylization with Style Matching ScoreYuxin Jiang, Liming Jiang, Shuai Yang, Jia-Wei Liu 等ICCV 2025 · 被引用 1 次
- PRINTER: Deformation-Aware Adversarial Learning for Virtual IHC Staining with In Situ FidelityYizhe Yuan, Bingsen Xue, Bangzheng Pu, Chengxiang Wang 等ACM MM 2025 · 被引用 1 次
- Multi-Window Gabor Transform Network for Ground Penetrating Radar B-Scan Image ReconstructionHuabin Wang, Yu Yang, Xinran Zhong, Zilong LingAAAI 2026
它引用的顶会 Paper21
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 被引用 24,064 次
- Training Generative Adversarial Networks with Limited DataTero Karras, Miika Aittala, Janne Hellsten, Samuli Laine 等NeurIPS 2020 · 被引用 2,345 次
- Differentiable Augmentation for Data-Efficient GAN TrainingShengyu Zhao, Zhijian Liu, Ji Lin, Jun-Yan Zhu 等NeurIPS 2020 · 被引用 707 次
- StyleGAN-NADA: CLIP-guided domain adaptation of image generatorsRinon Gal, Or Patashnik, Haggai Maron, Amit H. Bermano 等SIGGRAPH 2022 · 被引用 501 次
相关 Paper
- Towards Counterfactual Image Manipulation via CLIPYingchen Yu, Fangneng Zhan, Rongliang Wu, Jiahui Zhang 等ACM MM 2022 · 被引用 33 次
- StyleCLIP: Text-Driven Manipulation of StyleGAN ImageryOr Patashnik, Zongze Wu, Eli Shechtman, Daniel Cohen-Or 等ICCV 2021 · 被引用 1,437 次
- AI Illustrator: Translating Raw Descriptions into Images by Prompt-based Cross-Modal GenerationYiyang Ma, Huan Yang, Bei Liu, Jianlong Fu 等ACM MM 2022 · 被引用 9 次
- Learning to generate line drawings that convey geometry and semanticsCaroline Chan, Frédo Durand, Phillip IsolaCVPR 2022 · 被引用 86 次
- Unpaired Cartoon Image Synthesis via Gated Cycle MappingYifang Men, Yuan Yao, Miaomiao Cui, Zhouhui Lian 等CVPR 2022 · 被引用 19 次
