Inversion-based Style Transfer with Diffusion Models
Yuxin Zhang, Nisha Huang, Fan Tang, Haibin Huang, Chongyang Ma, Weiming Dong, Changsheng Xu
摘要
The artistic style within a painting is the means of expression, which includes not only the painting material, colors, and brushstrokes, but also the high-level attributes, including semantic elements and object shapes. Previous arbitrary example-guided artistic image generation methods often fail to control shape changes or convey elements. Pre-trained text-to-image synthesis diffusion probabilistic models have achieved remarkable quality but often require extensive textual descriptions to accurately portray the attributes of a particular painting. The uniqueness of an artwork lies in the fact that it cannot be adequately explained with normal language. Our key idea is to learn the artistic style directly from a single painting and then guide the synthesis without providing complex textual descriptions. Specifically, we perceive style as a learnable textual description of a painting. We propose an inversion-based style transfer method (InST), which can efficiently and accurately learn the key information of an image, thus capturing and transferring the artistic style of a painting. We demonstrate the quality and efficiency of our method on numerous paintings of various artists and styles. Codes are available at https://github.com/zyxElsa/InST.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper113
- CSGO: Content-Style Composition in Text-to-Image GenerationPeng Xing, Haofan Wang, Yanpeng Sun, Qixun Wang 等NeurIPS 2025 · 被引用 94 次
- FocalDreamer: Text-Driven 3D Editing via Focal-Fusion AssemblyYuhan Li, Yishun Dou, Yue Shi, Yu Lei 等AAAI 2024 · 被引用 91 次
- General Image-to-Image Translation with One-Shot Image GuidanceBin Cheng, Zuhao Liu, Yunbo Peng, Yue LinICCV 2023 · 被引用 60 次
- S2WAT: Image Style Transfer via Hierarchical Vision Transformer Using Strips Window AttentionChiyu Zhang, Xiaogang Xu, Lei Wang, Zaiyan Dai 等AAAI 2024 · 被引用 58 次
- DEADiff: An Efficient Stylization Diffusion Model with Disentangled RepresentationsTianhao Qi, Shancheng Fang, Yanze Wu, Hongtao Xie 等CVPR 2024 · 被引用 57 次
它引用的顶会 Paper28
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- Diffusion Models Beat GANs on Image SynthesisPrafulla Dhariwal, Alexander Quinn NicholNeurIPS 2021 · 被引用 13,211 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- Denoising Diffusion Implicit ModelsJiaming Song, Chenlin Meng, Stefano ErmonICLR 2021 · 被引用 11,743 次
相关 Paper
- SigStyle: Signature Style Transfer via Personalized Text-to-Image ModelsYe Wang, Tongyuan Bai, Xuping Xie, Zili Yi 等AAAI 2025 · 被引用 5 次
- Diverse Image Style Transfer via Invertible Cross-Space MappingHaibo Chen, Lei Zhao, Huiming Zhang, Zhizhong Wang 等ICCV 2021 · 被引用 42 次
- DualAST: Dual Style-Learning Networks for Artistic Style TransferHaibo Chen, Lei Zhao, Zhizhong Wang, Huiming Zhang 等CVPR 2021
- Style Injection in Diffusion: A Training-Free Approach for Adapting Large-Scale Diffusion Models for Style TransferJiwoo Chung, Sangeek Hyun, Jae-Pil HeoCVPR 2024
- StylerDALLE: Language-Guided Style Transfer Using a Vector-Quantized Tokenizer of a Large-Scale Generative ModelZipeng Xu, Enver Sangineto, Nicu SebeICCV 2023 · 被引用 16 次
