AesCLIP: Multi-Attribute Contrastive Learning for Image Aesthetics Assessment
Xiangfei Sheng, Leida Li, Pengfei Chen, Jinjian Wu, Weisheng Dong, Yuzhe Yang, Liwu Xu, Yaqian Li, Guangming Shi
摘要
Image aesthetics assessment (IAA) aims at predicting the aesthetic quality of images. Recently, large pre-trained vision-language models, like CLIP, have shown impressive performances on various visual tasks. When it comes to IAA, a straightforward way is to finetune the CLIP image encoder using aesthetic images. However, this can only achieve limited success without considering the uniqueness of multimodal data in the aesthetics domain. People usually assess image aesthetics according to fine-grained visual attributes, e.g., color, light and composition. However, how to learn aesthetics-aware attributes from CLIP-based semantic space has not been addressed before. With this motivation, this paper presents a CLIP-based multi-attribute contrastive learning framework for IAA, dubbed AesCLIP. Specifically, AesCLIP consists of two major components, i.e., aesthetic attribute-based comment classification and attribute-aware learning. The former classifies the aesthetic comments into different attribute categories. Then the latter learns an aesthetic attribute-aware representation by contrastive learning, aiming to mitigate the domain shift from the general visual domain to the aesthetics domain. Extensive experiments have been done by using the pre-trained AesCLIP on four popular IAA databases, and the results demonstrate the advantage of AesCLIP over the state-of-the-arts. The source code will be public at https://github.com/OPPOMKLab/AesCLIP.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
引用它的顶会 Paper6
- AesExpert: Towards Multi-modality Foundation Model for Image Aesthetics PerceptionYipo Huang, Xiangfei Sheng, Zhichao Yang, Quan Yuan 等ACM MM 2024 · 被引用 34 次
- Fine-grained Image Aesthetic Assessment: Learning Discriminative Scores from Relative RanksZhichao Yang, Jianjie Wang, Zhixianhe Zhang, Pangu Xie 等CVPR 2026 · 被引用 5 次
- Fine-grained Image Quality Assessment for Perceptual Image RestorationXiangfei Sheng, Xiaofeng Pan, Zhichao Yang, Pengfei Chen 等AAAI 2026 · 被引用 4 次
- FPEM: Face Prior Enhanced Facial Attractiveness Prediction for Live Videos with Face RetouchingHui Li, Xiaoyu Ren, Hongjiu Yu, Ying Chen 等ICCV 2025 · 被引用 1 次
- Charm: The Missing Piece in ViT Fine-Tuning for Image Aesthetic AssessmentFatemeh Behrad, Tinne Tuytelaars, Johan WagemansCVPR 2025
相关 Paper
- Exploring CLIP for Assessing the Look and Feel of ImagesJianyi Wang, Kelvin C. K. Chan, Chen Change LoyAAAI 2023 · 被引用 1,208 次
- VILA: Learning Image Aesthetics from User Comments with Vision-Language PretrainingJunjie Ke, Keren Ye, Jiahui Yu, Yonghui Wu 等CVPR 2023
- Semantics-Aware Image Aesthetics Assessment using Tag Matching and Contrastive RankingZhichao Yang, Leida Li, Pengfei Chen, Jinjian Wu 等ACM MM 2024 · 被引用 6 次
- Attribute-Driven Multimodal Hierarchical Prompts for Image Aesthetic Quality AssessmentHancheng Zhu, Ju Shi, Zhiwen Shao, Rui Yao 等ACM MM 2024 · 被引用 8 次
- CRIS: CLIP-Driven Referring Image SegmentationZhaoqing Wang, Yu Lu, Qiang Li, Xunqiang Tao 等CVPR 2022 · 被引用 337 次
