Paint Transformer: Feed Forward Neural Painting with Stroke Prediction
Songhua Liu, Tianwei Lin, Dongliang He, Fu Li, Ruifeng Deng, Xin Li, Errui Ding, Hao Wang
摘要
Neural painting refers to the procedure of producing a series of strokes for a given image and non-photo-realistically recreating it using neural networks. While reinforcement learning (RL) based agents can generate a stroke sequence step by step for this task, it is not easy to train a stable RL agent. On the other hand, stroke optimization methods search for a set of stroke parameters iteratively in a large search space; such low efficiency significantly limits their prevalence and practicality. Different from previous methods, in this paper, we formulate the task as a set prediction problem and propose a novel Transformer-based framework, dubbed Paint Transformer, to predict the parameters of a stroke set with a feed forward network. This way, our model can generate a set of strokes in parallel and obtain the final painting of size 512 × 512 in near real time. More importantly, since there is no dataset available for training the Paint Transformer, we devise a self-training pipeline such that it can be trained without any off-the-shelf dataset while still achieving excellent generalization capability. Experiments demonstrate that our method achieves better painting performance than previous ones with cheaper training and inference costs. Codes and models are available on https://github.com/wzmsltw/PaintTransformer.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper21
- Learning to generate line drawings that convey geometry and semanticsCaroline Chan, Frédo Durand, Phillip IsolaCVPR 2022 · 被引用 86 次
- Towards Layer-wise Image VectorizationXu Ma, Yuqian Zhou, Xingqian Xu, Bin Sun 等CVPR 2022 · 被引用 56 次
- Stroke-based Neural Painting and Stylization with Dynamically Predicted Painting RegionTeng Hu, Ran Yi, Haokun Zhu, Liang Liu 等ACM MM 2023 · 被引用 23 次
- Collaborative Transformers for Grounded Situation RecognitionJunhyeong Cho, Youngseok Yoon, Suha KwakCVPR 2022 · 被引用 23 次
- Optimize & Reduce: A Top-Down Approach for Image VectorizationOr Hirschorn, Amir Jevnisek, Shai AvidanAAAI 2024 · 被引用 20 次
它引用的顶会 Paper8
- FCOS: Fully Convolutional One-Stage Object DetectionZhi Tian, Chunhua Shen, Hao Chen, Tong HeICCV 2019 · 被引用 6,042 次
- Rethinking Rotated Object Detection with Gaussian Wasserstein Distance LossXue Yang, Junchi Yan, Qi Ming, Wentao Wang 等ICML 2021 · 被引用 572 次
- AdaAttN: Revisit Attention Mechanism in Arbitrary Neural Style TransferSonghua Liu, Tianwei Lin, Dongliang He, Fu Li 等ICCV 2021 · 被引用 421 次
- Learning to Paint With Model-Based Deep Reinforcement LearningZhewei Huang, Shuchang Zhou, Wen HengICCV 2019 · 被引用 180 次
- Drafting and Revision: Laplacian Pyramid Network for Fast High-Quality Artistic Style TransferTianwei Lin, Zhuoqi Ma, Fu Li, Dongliang He 等CVPR 2021
相关 Paper
- Stylized Neural PaintingZhengxia Zou, Tianyang Shi, Shuang Qiu, Yi Yuan 等CVPR 2021
- Combining Semantic Guidance and Deep Reinforcement Learning for Generating Human Level PaintingsJaskirat Singh, Liang ZhengCVPR 2021
- RenderFormer: Transformer-based Neural Rendering of Triangle Meshes with Global IlluminationChong Zeng, Yue Dong, Pieter Peers, Hongzhi Wu 等SIGGRAPH 2025 · 被引用 5 次
- Rethinking Style Transfer: From Pixels to Parameterized BrushstrokesDmytro Kotovenko, Matthias Wright, Arthur Heimbrecht, Björn OmmerCVPR 2021
- Painting Many Pasts: Synthesizing Time Lapse Videos of PaintingsAmy Zhao, Guha Balakrishnan, Kathleen M. Lewis, Frédo Durand 等CVPR 2020
