Domain Enhanced Arbitrary Image Style Transfer via Contrastive Learning
Yuxin Zhang, Fan Tang, Weiming Dong, Haibin Huang, Chongyang Ma, Tong-Yee Lee, Changsheng Xu
Abstract
In this work, we tackle the challenging problem of arbitrary image style transfer using a novel style feature representation learning method. A suitable style representation, as a key component in image stylization tasks, is essential to achieve satisfactory results. Existing deep neural network based approaches achieve reasonable results with the guidance from second-order statistics such as Gram matrix of content features. However, they do not leverage sufficient style information, which results in artifacts such as local distortions and style inconsistency. To address these issues, we propose to learn style representation directly from image features instead of their second-order statistics, by analyzing the similarities and differences between multiple styles and considering the style distribution. Specifically, we present Contrastive Arbitrary Style Transfer (CAST), which is a new style representation learning and style transfer method via contrastive learning. Our framework consists of three key components, i.e., a multi-layer style projector for style code encoding, a domain enhancement module for effective learning of style distribution, and a generative network for image style transfer. We conduct qualitative and quantitative evaluations comprehensively to demonstrate that our approach achieves significantly better results compared to those obtained via state-of-the-art methods. Code and models are available at https://github.com/zyxElsa/CAST_pytorch.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext a6bddc51-eb9e-4d12-9985-2459389f57c3Cited by top-tier papers31
- StyleDiffusion: Controllable Disentangled Style Transfer via Diffusion ModelsZhizhong Wang, Lei Zhao, Wei XingICCV 2023 · 219 citations
- FontDiffuser: One-Shot Font Generation via Denoising Diffusion with Multi-Scale Content Aggregation and Style Contrastive LearningZhenhua Yang, Dezhi Peng, Yuxin Kong, Yuyi Zhang et al.AAAI 2024 · 90 citations
- AesPA-Net: Aesthetic Pattern-Aware Style Transfer NetworksKibeom Hong, Seogkyu Jeon, Junsoo Lee, Namhyuk Ahn et al.ICCV 2023 · 69 citations
- General Image-to-Image Translation with One-Shot Image GuidanceBin Cheng, Zuhao Liu, Yunbo Peng, Yue LinICCV 2023 · 60 citations
- S2WAT: Image Style Transfer via Hierarchical Vision Transformer Using Strips Window AttentionChiyu Zhang, Xiaogang Xu, Lei Wang, Zaiyan Dai et al.AAAI 2024 · 58 citations
Builds on12
- StyTr2: Image Style Transfer with TransformersYingying Deng, Fan Tang, Weiming Dong, Chongyang Ma et al.CVPR 2022 · 345 citations
- ContraGAN: Contrastive Learning for Conditional Image GenerationMinguk Kang, Jaesik ParkNeurIPS 2020 · 216 citations
- Dynamic Instance Normalization for Arbitrary Style TransferYongcheng Jing, Xiao Liu, Yukang Ding, Xinchao Wang et al.AAAI 2020 · 212 citations
- Arbitrary Video Style Transfer via Multi-Channel CorrelationYingying Deng, Fan Tang, Weiming Dong, Haibin Huang et al.AAAI 2021 · 197 citations
- Arbitrary Style Transfer via Multi-Adaptation NetworkYingying Deng, Fan Tang, Weiming Dong, Wen Sun et al.ACM MM 2020 · 194 citations
Related papers
- Artistic Style Transfer with Internal-external Learning and Contrastive LearningHaibo Chen, Lei Zhao, Zhizhong Wang, Huiming Zhang et al.NeurIPS 2021 · 243 citations
- Dual-head Genre-instance Transformer Network for Arbitrary Style TransferMeichen Liu, Shuting He, Songnan Lin, Bihan WenACM MM 2024 · 3 citations
- Diversified Arbitrary Style Transfer via Deep Feature PerturbationZhizhong Wang, Lei Zhao, Haibo Chen, Lihong Qiu et al.CVPR 2020
- AesUST: Towards Aesthetic-Enhanced Universal Style TransferZhizhong Wang, Zhanjie Zhang, Lei Zhao, Zhiwen Zuo et al.ACM MM 2022 · 70 citations
- Multimodal Style Transfer via Graph CutsYulun Zhang, Chen Fang, Yilin Wang, Zhaowen Wang et al.ICCV 2019 · 92 citations
