A³: Towards Advertising Aesthetic Assessment
Kaiyuan Ji, Yixuan Gao, Lu Sun, Yushuo Zheng, Zijian Chen, Jianbo Zhang, Xiangyang Zhu, Yuan Tian, Zicheng Zhang, Guangtao Zhai
Abstract
Advertising images significantly impact commercial conversion rates and brand equity, yet current evaluation methods rely on subjective judgments, lacking scalability, standardized criteria, and interpretability. To address these challenges, we present A 3 (Advertising Aesthetic Assessment), a comprehensive framework encompassing four components: a paradigm (A 3 -Law), a dataset (A 3 -Dataset), a multimodal large language model (A 3 -Align), and a benchmark (A 3 -Bench). Central to A 3 is a theorydriven paradigm, A 3 -Law, comprising three hierarchical stages: (1) Perceptual Attention, evaluating perceptual image signals for their ability to attract attention; (2) Formal Interest, assessing formal composition of image color and spatial layout in evoking interest; and (3) Desire Impact, measuring desire evocation from images and their persuasive impact. Building on A 3 -Law, we construct A 3 -Dataset with 120K instruction-response pairs from 30K advertising images, each richly annotated with multi-dimensional labels and Chain-of-Thought (CoT) rationales. We further develop A 3 -Align, trained under A 3 -Law with CoT-guided learning on A 3 -Dataset. Extensive experiments on A 3 -Bench demonstrate that A 3 -Align achieves superior alignment with A 3 -Law compared to existing models, and this alignment generalizes well to quality advertisement selection and prescriptive advertisement critique, indicating its potential for broader deployment. Dataset, code, and models can be found at: https://github.com/euleryuan/A3-Align
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 603e3de7-f08c-4472-8cdc-2cbf81d60964Builds on15
- Language Models Don't Always Say What They Think: Unfaithful Explanations in Chain-of-Thought PromptingMiles Turpin, Julian Michael, Ethan Perez, Samuel R. BowmanNeurIPS 2023 · 1,792 citations
- MMMU: A Massive Multi-Discipline Multimodal Understanding and Reasoning Benchmark for Expert AGIXiang Yue, Yuansheng Ni, Tianyu Zheng, Kai Zhang et al.CVPR 2024 · 213 citations
- Accelerating Non-Maximum Suppression: A Graph Theory PerspectiveKing-Siong Si, Lu Sun, Weizhan Zhang, Tieliang Gong et al.NeurIPS 2024 · 14 citations
- Generalizable Video Quality Assessment via Weak-to-Strong LearningLinhan Cao, Wei Sun, Xiangyang Zhu, Kaiwei Zhang et al.CVPR 2026 · 9 citations
- Image Quality Assessment for Embodied AIChunyi Li, Jiahao Xiao, Jianbo Zhang, Farong Wen et al.ICLR 2026 · 3 citations
Related papers
- Bridging Cognitive Gap: Hierarchical Description Learning for Artistic Image Aesthetics AssessmentHenglin Liu, Nisha Huang, Chang Liu, Jiangpeng Yan et al.AAAI 2026 · 1 citation
- InstructCrop: Teaching Multimodal Large Language Models to Crop Aesthetic ImagesXiangfei Sheng, Pangu Xie, Weidong Zou, Pengfei Chen et al.ACM MM 2025
- Aligning Vision Models with Human Aesthetics in Retrieval: Benchmarks and AlgorithmsMiaosen Zhang, Yixuan Wei, Zhen Xing, Yifei Ma et al.NeurIPS 2024 · 2 citations
- Can Vision-Language Models Assess Graphic Design Aesthetics? A Benchmark, Evaluation, and Dataset PerspectiveRuichuan An, Shizhao Sun, Danqing Huang, Mingxi Cheng et al.ICLR 2026 · 6 citations
- ArtiMuse: Fine-Grained Image Aesthetics Assessment with Joint Scoring and Expert-Level UnderstandingShuo Cao, Nan Ma, Jiayang Li, Xiaohui Li et al.CVPR 2026 · 38 citations
