APISR: Anime Production Inspired Real-World Anime Super-Resolution
Boyang Wang, Fengyu Yang, Xihang Yu, Chao Zhang, Hanbin Zhao
Abstract
While real-world anime super-resolution (SR) has gained increasing attention in the SR community, existing methods still adopt techniques from the photorealistic domain. In this paper, we analyze the anime production workflow and rethink how to use characteristics of it for the sake of the real-world anime SR. First, we argue that video networks and datasets are not necessary for anime SR due to the repetition use of hand-drawing frames. Instead, we propose an anime image collection pipeline by choosing the least compressed and the most informative frames from the video sources. Based on this pipeline, we introduce the Anime Production-oriented Image (API) dataset. In addition, we identify two anime-specific challenges of distorted and faint hand-drawn lines and unwanted color artifacts. We address the first issue by introducing a prediction-oriented compression module in the image degradation model and a pseudo-ground truth preparation with enhanced hand-drawn lines. In addition, we introduce the balanced twin perceptual loss combining both anime and photorealistic high-level features to mitigate unwanted color artifacts and increase visual clarity. We evaluate our method through extensive experiments on the public benchmark, showing our method outperforms state-of-the-art anime dataset-trained approaches. The code is available at https://github.com/Kiteretsu77/APISR.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers8
- How to Continually Adapt Text-to-Image Diffusion Models for Flexible Customization?Jiahua Dong, Wenqi Liang, Hongliu Li, Duzhen Zhang et al.NeurIPS 2024 · 42 citations
- BELM: Bidirectional Explicit Linear Multi-step Sampler for Exact Inversion in Diffusion ModelsFangyikang Wang, Hubery Yin, Yuejiang Dong, Huminhao Zhu et al.NeurIPS 2024 · 37 citations
- Frame In-N-Out: Unbounded Controllable Image-to-Video GenerationBoyang Wang, Xuweiyi Chen, Matheus Gadelha, Zezhou ChengNeurIPS 2025 · 9 citations
- UARE: A Unified Vision-Language Model for Image Quality Assessment, Restoration, and EnhancementWeiqi Li, Xuanyu Zhang, Bin Chen, Jingfen Xie et al.CVPR 2026 · 5 citations
- GenDR: Lighten Generative Detail RestorationYan Wang, Shijie Zhao, Kexin Zhang, Junlin Li et al.ICLR 2026 · 5 citations
Builds on17
- Adding Conditional Control to Text-to-Image Diffusion ModelsLvmin Zhang, Anyi Rao, Maneesh AgrawalaICCV 2023 · 6,759 citations
- AnimateDiff: Animate Your Personalized Text-to-Image Diffusion Models without Specific TuningYuwei Guo, Ceyuan Yang, Anyi Rao, Zhengyang Liang et al.ICLR 2024 · 1,493 citations
- Exploring CLIP for Assessing the Look and Feel of ImagesJianyi Wang, Kelvin C. K. Chan, Chen Change LoyAAAI 2023 · 1,208 citations
- Investigating Tradeoffs in Real-World Video Super-ResolutionKelvin C. K. Chan, Shangchen Zhou, Xiangyu Xu, Chen Change LoyCVPR 2022 · 106 citations
- Space-Time Video Super-Resolution Using Temporal ProfilesZeyu Xiao, Zhiwei Xiong, Xueyang Fu, Dong Liu et al.ACM MM 2020 · 54 citations
Related papers
- Learning Data-Driven Vector-Quantized Degradation Model for Animation Video Super-ResolutionZixi Tuo, Huan Yang, Jianlong Fu, Yujie Dun et al.ICCV 2023 · 5 citations
- AnimeSR: Learning Real-World Super-Resolution Models for Animation VideosYanze Wu, Xintao Wang, Gen Li, Ying ShanNeurIPS 2022 · 46 citations
- Deep Geometrized Cartoon Line InbetweeningLi Siyao, Tianpei Gu, Weiye Xiao, Henghui Ding et al.ICCV 2023 · 16 citations
- Bridging the Gap: Sketch-Aware Interpolation Network for High-Quality Animation Sketch InbetweeningJiaming Shen, Kun Hu, Wei Bao, Chang Wen Chen et al.ACM MM 2024 · 5 citations
- Scenimefy: Learning to Craft Anime Scene via Semi-Supervised Image-to-Image TranslationYuxin Jiang, Liming Jiang, Shuai Yang, Chen Change LoyICCV 2023 · 25 citations
