PosterCraft: Rethinking High-Quality Aesthetic Poster Generation in a Unified Framework
Sixiang Chen, Jianyu Lai, Jialin Gao, Tian Ye, Haoyu Chen, Hengyu Shi, Shitong Shao, Yunlong Lin, Song Fei, Zhaohu Xing, Yeying Jin, Junfeng Luo
摘要
Generating aesthetic posters is more challenging than simple design images: it requires not only precise text rendering but also the seamless integration of abstract artistic content, striking layouts, and overall stylistic harmony. To address this, we propose PosterCraft, a unified framework that abandons prior modular pipelines and rigid, predefined layouts, allowing the model to freely explore coherent, visually compelling compositions. PosterCraft employs a carefully designed, cascaded workflow to optimize the generation of high-aesthetic posters: (i) large-scale text-rendering optimization on our newly introduced Text-Render-2M dataset; (ii) region-aware supervised fine-tuning on HQ-Poster100K; (iii) aesthetic-text-reinforcement learning via best-of-n preference optimization; and (iv) joint vision-language feedback refinement. Each stage is supported by a fully automated data-construction pipeline tailored to its specific needs, enabling robust training without complex architectural modifications. Evaluated on multiple experiments, PosterCraft significantly outperforms open-source baselines in rendering accuracy, layout coherence, and overall visual appeal-approaching the quality of SOTA commercial systems. Our code, models, and datasets can be found in the Project page: https://ephemeral182.github.io/PosterCraft
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper16
- FLUX-Reason-6M & PRISM-Bench: A Million-Scale Text-to-Image Reasoning Dataset and Comprehensive BenchmarkRongyao Fang, Aldrich Yu, Chengqi Duan, Linjiang Huang 等ICLR 2026 · 被引用 37 次
- OpenGPT-4o-Image: A Comprehensive Dataset for Advanced Image Generation and Editingzhihong Chen, Xuehai Bai, Yang Shi, Chaoyou Fu 等ICML 2026 · 被引用 24 次
- PosterOmni: Generalized Artistic Poster Creation via Task Distillation and Unified Reward FeedbackSixiang Chen, Jianyu LAI, Jialin Gao, Hengyu Shi 等CVPR 2026 · 被引用 9 次
- PosterReward: Unlocking Accurate Evaluation for High-Quality Graphic Design GenerationJianyu LAI, Sixiang Chen, Jialin Gao, Hengyu Shi 等CVPR 2026 · 被引用 5 次
- InnoAds-Composer: Efficient Condition Composition for E-Commerce Poster GenerationYuxin Qin, Ke Cao, Haowei Liu, Ao Ma 等CVPR 2026 · 被引用 5 次
它引用的顶会 Paper21
- Direct Preference Optimization: Your Language Model is Secretly a Reward ModelRafael Rafailov, Archit Sharma, Eric Mitchell, Christopher D. Manning 等NeurIPS 2023 · 被引用 10,924 次
- SDXL: Improving Latent Diffusion Models for High-Resolution Image SynthesisDustin Podell, Zion English, Kyle Lacey, Andreas Blattmann 等ICLR 2024 · 被引用 4,569 次
- LayoutGPT: Compositional Visual Planning and Generation with Large Language ModelsWeixi Feng, Wanrong Zhu, Tsu-Jui Fu, Varun Jampani 等NeurIPS 2023 · 被引用 462 次
- TextDiffuser: Diffusion Models as Text PaintersJingye Chen, Yupan Huang, Tengchao Lv, Lei Cui 等NeurIPS 2023 · 被引用 290 次
- Mastering Text-to-Image Diffusion: Recaptioning, Planning, and Generating with Multimodal LLMsLing Yang, Zhaochen Yu, Chenlin Meng, Minkai Xu 等ICML 2024 · 被引用 231 次
相关 Paper
- PosterAgent: Agentic Poster Generation via Stage-Aware Reinforcement LearningZhuocheng Yu, Feng Zhang, Sujian Li, Kai JiaICML 2026
- POSTA: A Go-to Framework for Customized Artistic Poster GenerationHaoyu Chen, Xiaojie Xu, Wenbo Li, Jingjing Ren 等CVPR 2025
- PosterVerse: A Full-Workflow Framework for Commercial-Grade Poster Generation with HTML-Based Scalable TypographyJunle Liu, Peirong Zhang, Yuyi Zhang, Pengyu Yan 等AAAI 2026 · 被引用 3 次
- AutoPP: Towards Automated Product Poster Generation and OptimizationJiahao Fan, Yuxin Qin, Wei Feng, Yanyin Chen 等AAAI 2026 · 被引用 2 次
- TextPainter: Multimodal Text Image Generation with Visual-harmony and Text-comprehension for Poster DesignYifan Gao, Jinpeng Lin, Min Zhou, Chuanbin Liu 等ACM MM 2023 · 被引用 6 次
