Style Injection in Diffusion: A Training-Free Approach for Adapting Large-Scale Diffusion Models for Style Transfer
Jiwoo Chung, Sangeek Hyun, Jae-Pil Heo
Abstract
Despite the impressive generative capabilities of diffusion models, existing diffusion model-based style transfer methods require inference-stage optimization (e.g. fine-tuning or textual inversion of style) which is timeconsuming, or fails to leverage the generative ability of large-scale diffusion models. To address these issues, we introduce a novel artistic style transfer method based on a pre-trained large-scale diffusion model without any optimization. Specifically, we manipulate the features of selfattention layers as the way the cross-attention mechanism works; in the generation process, substituting the key and value of content with those of style image. This approach provides several desirable characteristics for style transfer including 1) preservation of content by transferring similar styles into similar image patches and 2) transfer of style based on similarity of local texture (e.g. edge) between content and style images. Furthermore, we introduce query preservation and attention temperature scaling to mitigate the issue of disruption of original content, and initial latent Adaptive Instance Normalization (AdaIN) to deal with the disharmonious color (failure to transfer the colors of style). Our experimental results demonstrate that our proposed method surpasses state-of-the-art methods in both conventional and diffusion-based style transfer baselines. Codes are available at github.com/jiwoogit/StyleID.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers86
- CSGO: Content-Style Composition in Text-to-Image GenerationPeng Xing, Haofan Wang, Yanpeng Sun, Qixun Wang et al.NeurIPS 2025 · 94 citations
- Unified In-Context Video EditingZixuan Ye, Xuanhua He, Quande Liu, Qiulin Wang et al.ICLR 2026 · 37 citations
- AsyncDiff: Parallelizing Diffusion Models by Asynchronous DenoisingZigeng Chen, Xinyin Ma, Gongfan Fang, Zhenxiong Tan et al.NeurIPS 2024 · 33 citations
- ThermalGen: Style-Disentangled Flow-Based Generative Models for RGB-to-Thermal Image TranslationJiuhong Xiao, Roshan Nayak, Ning Zhang, Daniel Tortei et al.NeurIPS 2025 · 18 citations
- HarmonyCut: Supporting Creative Chinese Paper-cutting Design with Form and Connotation HarmonyHuanchen Wang, Tianrun Qiu, Jiaping Li, Zhicong Lu et al.CHI 2025 · 17 citations
Builds on34
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- Denoising Diffusion Implicit ModelsJiaming Song, Chenlin Meng, Stefano ErmonICLR 2021 · 11,743 citations
- Photorealistic Text-to-Image Diffusion Models with Deep Language UnderstandingChitwan Saharia, William Chan, Saurabh Saxena, Lala Li et al.NeurIPS 2022 · 8,965 citations
- BLIP: Bootstrapping Language-Image Pre-training for Unified Vision-Language Understanding and GenerationJunnan Li, Dongxu Li, Caiming Xiong, Steven C. H. HoiICML 2022 · 6,549 citations
Related papers
- Inversion-based Style Transfer with Diffusion ModelsYuxin Zhang, Nisha Huang, Fan Tang, Haibin Huang et al.CVPR 2023
- FreePIH: Training-Free Painterly Image Harmonization with Diffusion ModelRuibin Li, Jingcai Guo, Qihua Zhou, Song GuoACM MM 2024 · 2 citations
- Attention Distillation: A Unified Approach to Visual Characteristics TransferYang Zhou, Xu Gao, Zichong Chen, Hui HuangCVPR 2025
- ArtBank: Artistic Style Transfer with Pre-trained Diffusion Model and Implicit Style Prompt BankZhanjie Zhang, Quanwei Zhang, Wei Xing, Guangyuan Li et al.AAAI 2024 · 32 citations
- ACID-Style: An Adaptive Condition Injection Diffusion Model for Arbitrary Style TransferTing Yang, Siyu Yang, Xiyao Liu, Songtao Wu et al.AAAI 2026
