PPJudge: Towards Human-Aligned Assessment of Artistic Painting Process
Shiqi Jiang, Xinpeng Li, Xi Mao, Changbo Wang, Chenhui Li
Abstract
Artistic image assessment has become a prominent research area in computer vision. In recent years, the field has witnessed a proliferation of datasets and methods designed to evaluate the aesthetic quality of paintings. However, most existing approaches focus solely on static final images, overlooking the dynamic and multi-stage nature of the artistic painting process. To address this gap, we propose a novel framework for human-aligned assessment of painting processes. Specifically, we introduce the Painting Process Assessment Dataset (PPAD)-the first large-scale dataset comprising real and synthetic painting process images, annotated by domain experts across eight detailed attributes. Furthermore, we present PPJudge (Painting Process Judge), a Transformer-based model enhanced with temporally-aware positional encoding and a heterogeneous mixture-of-experts architecture, enabling effective assessment of the painting process. Experimental results demonstrate that our method outperforms existing baselines in accuracy, robustness, and alignment with human judgment, offering new insights into computational creativity and art education.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 42b7741e-6bf3-441c-845d-0349f70f5c95Cited by top-tier papers1
Ask how each one uses itBuilds on16
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu et al.ICCV 2021 · 31,683 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- SlowFast Networks for Video RecognitionChristoph Feichtenhofer, Haoqi Fan, Jitendra Malik, Kaiming HeICCV 2019 · 4,104 citations
- ViViT: A Video Vision TransformerAnurag Arnab, Mostafa Dehghani, Georg Heigold, Chen Sun et al.ICCV 2021 · 2,947 citations
Related papers
- Towards Artistic Image Aesthetics Assessment: a Large-scale Dataset and a New MethodRan Yi, Haoyuan Tian, Zhihao Gu, Yu-Kun Lai et al.CVPR 2023
- AACP: Aesthetics Assessment of Children's Paintings Based on Self-Supervised LearningShiqi Jiang, Ning Li, Chen Shi, Liping Guo et al.AAAI 2024 · 2 citations
- Assessing Eye Aesthetics for Automatic Multi-Reference Eye In-PaintingBo Yan, Qing Lin, Weimin Tan, Shili ZhouCVPR 2020
- Bridging Cognitive Gap: Hierarchical Description Learning for Artistic Image Aesthetics AssessmentHenglin Liu, Nisha Huang, Chang Liu, Jiangpeng Yan et al.AAAI 2026 · 1 citation
- Thinking Image Color Aesthetics Assessment: Models, Datasets and BenchmarksShuai He, Anlong Ming, Yaqi Li, Jinyuan Sun et al.ICCV 2023 · 36 citations
