TextPainter: Multimodal Text Image Generation with Visual-harmony and Text-comprehension for Poster Design
Yifan Gao, Jinpeng Lin, Min Zhou, Chuanbin Liu, Hongtao Xie, Tiezheng Ge, Yuning Jiang
摘要
Text design is one of the most critical procedures in poster design, as it relies heavily on the creativity and expertise of humans to design text images considering the visual harmony and text-semantic. This study introduces TextPainter, a novel multimodal approach that leverages contextual visual information and corresponding text semantics to generate text images. Specifically, TextPainter takes the global-local background image as a hint of style and guides the text image generation with visual harmony. Furthermore, we leverage the language model and introduce a text comprehension module to achieve both sentence-level and word-level style variations. Besides, we construct the PosterT80K dataset, consisting of about 80K posters annotated with sentence-level bounding boxes and text contents. We hope this dataset will pave the way for further research on multimodal text image generation. Extensive quantitative and qualitative experiments demonstrate that TextPainter can generate visually-and-semantically-harmonious text images for posters.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper7
- InnoAds-Composer: Efficient Condition Composition for E-Commerce Poster GenerationYuxin Qin, Ke Cao, Haowei Liu, Ao Ma 等CVPR 2026 · 被引用 5 次
- Rethinking Layered Graphic Design Generation with a Top-Down ApproachJingye Chen, Zhaowen Wang, Nanxuan Zhao, Li Zhang 等ICCV 2025 · 被引用 4 次
- OmniText: A Training-Free Generalist for Controllable Text-Image ManipulationAgus Gunawan, Samuel Teodoro, Yun Chen, Soo Ye Kim 等ICLR 2026 · 被引用 3 次
- LOCATEdit: Graph Laplacian Optimized Cross Attention for Localized Text-Guided Image EditingAchint Soni, Meet Soni, Sirisha RambhatlaICCV 2025 · 被引用 1 次
- PSDesigner: Automated Graphic Design with a Human-Like Creative WorkflowXincheng Shuai, Song Tang, Yutong Huang, Henghui Ding 等CVPR 2026 · 被引用 1 次
它引用的顶会 Paper6
- Few-Shot Font Generation by Learning Fine-Grained Local StylesLicheng Tang, Yiyang Cai, Jiaming Liu, Zhibin Hong 等CVPR 2022 · 被引用 77 次
- GAN-Based Unpaired Chinese Character Image Translation via Skeleton Transformation and Stroke RenderingYiming Gao, Jiangqin WuAAAI 2020 · 被引用 71 次
- XMP-Font: Self-Supervised Cross-Modality Pre-training for Few-Shot Font GenerationWei Liu, Fangyue Liu, Fei Ding, Qian He 等CVPR 2022 · 被引用 64 次
- Geometry Aligned Variational Transformer for Image-conditioned Layout GenerationYunning Cao, Ye Ma, Min Zhou, Chuanbin Liu 等ACM MM 2022 · 被引用 37 次
- Self-Supervised Text Erasing with Controllable Image SynthesisGangwei Jiang, Shiyao Wang, Tiezheng Ge, Yuning Jiang 等ACM MM 2022 · 被引用 10 次
相关 Paper
- AutoPoster: A Highly Automatic and Content-aware Design System for Advertising Poster GenerationJinpeng Lin, Min Zhou, Ye Ma, Yifan Gao 等ACM MM 2023 · 被引用 24 次
- Prompt2Poster: Automatically Artistic Chinese Poster Creation from Prompt OnlyShaodong Wang, Yunyang Ge, Liuhan Chen, Haiyang Zhou 等ACM MM 2024 · 被引用 5 次
- POSTA: A Go-to Framework for Customized Artistic Poster GenerationHaoyu Chen, Xiaojie Xu, Wenbo Li, Jingjing Ren 等CVPR 2025
- PosterCraft: Rethinking High-Quality Aesthetic Poster Generation in a Unified FrameworkSixiang Chen, Jianyu Lai, Jialin Gao, Tian Ye 等ICLR 2026 · 被引用 43 次
- PosterReward: Unlocking Accurate Evaluation for High-Quality Graphic Design GenerationJianyu LAI, Sixiang Chen, Jialin Gao, Hengyu Shi 等CVPR 2026 · 被引用 5 次
