DreamStyler: Paint by Style Inversion with Text-to-Image Diffusion Models
Namhyuk Ahn, Junsoo Lee, Chunggi Lee, Kunhee Kim, Daesik Kim, Seung-Hun Nam, Kibeom Hong
摘要
Recent progresses in large-scale text-to-image models have yielded remarkable accomplishments, finding various applications in art domain. However, expressing unique characteristics of an artwork (e.g. brushwork, colortone, or composition) with text prompts alone may encounter limitations due to the inherent constraints of verbal description. To this end, we introduce DreamStyler, a novel framework designed for artistic image synthesis, proficient in both text-to-image synthesis and style transfer. DreamStyler optimizes a multi-stage textual embedding with a context-aware text prompt, resulting in prominent image quality. In addition, with content and style guidance, DreamStyler exhibits flexibility to accommodate a range of style references. Experimental results demonstrate its superior performance across multiple scenarios, suggesting its promising potential in artistic product creation. Project page: https://nmhkahn.github.io/dreamstyler/ .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper14
- The Chosen One: Consistent Characters in Text-to-Image Diffusion ModelsOmri Avrahami, Amir Hertz, Yael Vinker, Moab Arar 等SIGGRAPH 2024 · 被引用 26 次
- InstructG2I: Synthesizing Images from Multimodal Attributed GraphsBowen Jin, Ziqi Pang, Bingjun Guo, Yu-Xiong Wang 等NeurIPS 2024 · 被引用 15 次
- FineStyle: Fine-grained Controllable Style Personalization for Text-to-image ModelsGong Zhang, Kihyuk Sohn, Meera Hahn, Humphrey Shi 等NeurIPS 2024 · 被引用 15 次
- Free-Lunch Color-Texture Disentanglement for Stylized Image GenerationJiang Qin, Alexandra Gomez-Villa, Senmao Li, Shiqi Yang 等NeurIPS 2025 · 被引用 12 次
- ArtEditor: Learning Customized Instructional Image Editor From Few-Shot ExamplesShijie Huang, Yiren Song, Yuxuan Zhang, Hailong Guo 等ICCV 2025 · 被引用 2 次
它引用的顶会 Paper16
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- Photorealistic Text-to-Image Diffusion Models with Deep Language UnderstandingChitwan Saharia, William Chan, Saurabh Saxena, Lala Li 等NeurIPS 2022 · 被引用 8,965 次
- BLIP-2: Bootstrapping Language-Image Pre-training with Frozen Image Encoders and Large Language ModelsJunnan Li, Dongxu Li, Silvio Savarese, Steven C. H. HoiICML 2023 · 被引用 7,873 次
- Adding Conditional Control to Text-to-Image Diffusion ModelsLvmin Zhang, Anyi Rao, Maneesh AgrawalaICCV 2023 · 被引用 6,759 次
相关 Paper
- Inversion-based Style Transfer with Diffusion ModelsYuxin Zhang, Nisha Huang, Fan Tang, Haibin Huang 等CVPR 2023
- DreamIdentity: Enhanced Editability for Efficient Face-Identity Preserved Image GenerationZhuowei Chen, Shancheng Fang, Wei Liu, Qian He 等AAAI 2024 · 被引用 26 次
- DreamStyle: A Unified Framework for Video StylizationMengtian Li, Jinshu Chen, Songtao Zhao, Wanquan Feng 等CVPR 2026 · 被引用 4 次
- IPDreamer: Appearance-Controllable 3D Object Generation with Complex Image PromptsBohan Zeng, Shanglin Li, Yutang Feng, Ling Yang 等ICLR 2025
- DesignDiffusion: High-Quality Text-to-Design Image Generation with Diffusion ModelsZhendong Wang, Jianmin Bao, Shuyang Gu, Dong Chen 等CVPR 2025
