Consistent Video Style Transfer via Compound Regularization
Wenjing Wang, Jizheng Xu, Li Zhang, Yue Wang, Jiaying Liu
Abstract
Recently, neural style transfer has drawn many attentions and significant progresses have been made, especially for image style transfer. However, flexible and consistent style transfer for videos remains a challenging problem. Existing training strategies, either using a significant amount of video data with optical flows or introducing single-frame regularizers, have limited performance on real videos. In this paper, we propose a novel interpretation of temporal consistency, based on which we analyze the drawbacks of existing training strategies; and then derive a new compound regularization. Experimental results show that the proposed regularization can better balance the spatial and temporal performance, which supports our modeling. Combining with the new cost formula, we design a zero-shot video style transfer framework. Moreover, for better feature migration, we introduce a new module to dynamically adjust inter-channel distributions. Quantitative and qualitative results demonstrate the superiority of our method over other state-of-the-art style transfer methods. Our project is publicly available at: https://daooshee.github . io/CompoundVST/.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 5bd4af69-dede-4385-a819-a0df5df0c887Cited by top-tier papers14
- AdaAttN: Revisit Attention Mechanism in Arbitrary Neural Style TransferSonghua Liu, Tianwei Lin, Dongliang He, Fu Li et al.ICCV 2021 · 421 citations
- Arbitrary Video Style Transfer via Multi-Channel CorrelationYingying Deng, Fan Tang, Weiming Dong, Haibin Huang et al.AAAI 2021 · 197 citations
- Learning to Stylize Novel ViewsHsin-Ping Huang, Hung-Yu Tseng, Saurabh Saini, Maneesh Singh et al.ICCV 2021 · 98 citations
- AesPA-Net: Aesthetic Pattern-Aware Style Transfer NetworksKibeom Hong, Seogkyu Jeon, Junsoo Lee, Namhyuk Ahn et al.ICCV 2023 · 69 citations
- Video Demoiréing with Relation-Based Temporal ConsistencyPeng Dai, Xin Yu, Lan Ma, Baoheng Zhang et al.CVPR 2022 · 23 citations
Related papers
- Stable Video Style Transfer Based on Partial Convolution with Depth-Aware SupervisionSonghua Liu, Hao Wu, Shoutong Luo, Zhengxing SunACM MM 2020 · 6 citations
- Fresco: Spatial-Temporal Correspondence for Zero-Shot Video TranslationShuai Yang, Yifan Zhou, Ziwei Liu, Chen Change LoyCVPR 2024 · 17 citations
- Preserving Global and Local Temporal Consistency for Arbitrary Video Style TransferXinxiao Wu, Jialu ChenACM MM 2020 · 14 citations
- FreeViS: Training-free Video Stylization with Inconsistent ReferencesJiacong Xu, Yiqun Mei, Ke Zhang, Vishal M. PatelICLR 2026 · 7 citations
- Unsupervised Coherent Video Cartoonization with Perceptual Motion ConsistencyZhenhuan Liu, Liang Li, Huajie Jiang, Xin Jin et al.AAAI 2022 · 7 citations
