Unsupervised Coherent Video Cartoonization with Perceptual Motion Consistency
Zhenhuan Liu, Liang Li, Huajie Jiang, Xin Jin, Dandan Tu, Shuhui Wang, Zheng-Jun Zha
摘要
In recent years, creative content generations like style transfer and neural photo editing have attracted more and more attention. Among these, cartoonization of real-world scenes has promising applications in entertainment and industry. Different from image translations focusing on improving the style effect of generated images, video cartoonization has additional requirements on the temporal consistency. In this paper, we propose a spatially-adaptive semantic alignment framework with perceptual motion consistency for coherent video cartoonization in an unsupervised manner. The semantic alignment module is designed to restore deformation of semantic structure caused by spatial information lost in the encoder-decoder architecture. Furthermore, we devise the spatio-temporal correlative map as a style-independent, global-aware regularization on the perceptual motion consistency. Deriving from similarity measurement of high-level features in photo and cartoon frames, it captures global semantic information beyond raw pixel-value in optical flow. Besides, the similarity measurement disentangles temporal relationships from domain-specific style properties, which helps regularize the temporal consistency without hurting style effects of cartoon images. Qualitative and quantitative experiments demonstrate our method is able to generate highly stylistic and temporal consistent cartoon videos.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper8
- Arbitrary Video Style Transfer via Multi-Channel CorrelationYingying Deng, Fan Tang, Weiming Dong, Haibin Huang 等AAAI 2021 · 被引用 197 次
- Blind Video Temporal Consistency via Deep Video PriorChenyang Lei, Yazhou Xing, Qifeng ChenNeurIPS 2020 · 被引用 134 次
- Consistent Video Style Transfer via Compound RegularizationWenjing Wang, Jizheng Xu, Li Zhang, Yue Wang 等AAAI 2020 · 被引用 50 次
- Multimodal Structure-Consistent Image-to-Image TranslationChe-Tsung Lin, Yen-Yi Wu, Po-Hao Hsu, Shang-Hong LaiAAAI 2020 · 被引用 24 次
- Learning to Transfer: Unsupervised Domain Translation via Meta-LearningJianxin Lin, Yijun Wang, Zhibo Chen, Tianyu HeAAAI 2020 · 被引用 10 次
相关 Paper
- Preserving Global and Local Temporal Consistency for Arbitrary Video Style TransferXinxiao Wu, Jialu ChenACM MM 2020 · 被引用 14 次
- CADQ: Attribute-Consistent Face Cartoonization with Cross-modal Aligned and Deformable QuantizationYongjie Hu, Yifan Jiang, Ziyun Li, Fei Gao 等ACM MM 2025
- Learning Temporally and Semantically Consistent Unpaired Video-to-Video Translation through Pseudo-Supervision from Synthetic Optical FlowKaihong Wang, Kumar Akash, Teruhisa MisuAAAI 2022 · 被引用 16 次
- Cartoon-Flow: A Flow-Based Generative Adversarial Network for Arbitrary-Style Photo CartoonizationJieun Lee, Hyeonwoo Kim, Jonghwa Shim, Eenjun HwangACM MM 2022 · 被引用 14 次
- STRIVE: Scene Text Replacement In VideosVijay Kumar B. G, Jeyasri Subramanian, Varnith Chordia, Eugene Bart 等ICCV 2021 · 被引用 14 次
