Transformer-Based No-Reference Image Quality Assessment via Supervised Contrastive Learning
Jinsong Shi, Pan Gao, Jie Qin
摘要
Image Quality Assessment (IQA) has long been a research hotspot in the field of image processing, especially No-Reference Image Quality Assessment (NR-IQA). Due to the powerful feature extraction ability, existing Convolution Neural Network (CNN) and Transformers based NR-IQA methods have achieved considerable progress. However, they still exhibit limited capability when facing unknown authentic distortion datasets. To further improve NR-IQA performance, in this paper, a novel supervised contrastive learning (SCL) and Transformer-based NR-IQA model SaTQA is proposed. We first train a model on a large-scale synthetic dataset by SCL (no image subjective score is required) to extract degradation features of images with various distortion types and levels. To further extract distortion information from images, we propose a backbone network incorporating the Multi-Stream Block (MSB) by combining the CNN inductive bias and Transformer long-term dependence modeling capability. Finally, we propose the Patch Attention Block (PAB) to obtain the final distorted image quality score by fusing the degradation features learned from contrastive learning with the perceptual distortion information extracted by the backbone network. Experimental results on seven standard IQA datasets show that SaTQA outperforms the state-of-the-art methods for both synthetic and authentic datasets. Code is available at https://github.com/I2-Multimedia-Lab/SaTQA
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- Q-Norm: Robust Representation Learning via Quality-Adaptive NormalizationLanning Zhang, Ying Zhou, Fei Gao, Ziyun Li 等ICCV 2025 · 被引用 1 次
- QuARF: Quality-Adaptive Receptive Fields for Degraded Image PerceptionFei Gao, Ying Zhou, Ziyun Li, Wenwang Han 等AAAI 2025
它引用的顶会 Paper8
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 被引用 24,064 次
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- Supervised Contrastive LearningPrannay Khosla, Piotr Teterwak, Chen Wang, Aaron Sarna 等NeurIPS 2020 · 被引用 7,049 次
- MUSIQ: Multi-scale Image Quality TransformerJunjie Ke, Qifei Wang, Yilin Wang, Peyman Milanfar 等ICCV 2021 · 被引用 1,325 次
- Blindly Assess Image Quality in the Wild Guided by a Self-Adaptive Hyper NetworkShaolin Su, Qingsen Yan, Yu Zhu, Cheng Zhang 等CVPR 2020
相关 Paper
- Data-Efficient Image Quality Assessment with Attention-Panel DecoderGuanyi Qin, Runze Hu, Yutao Liu, Xiawu Zheng 等AAAI 2023 · 被引用 113 次
- Long Short-term Convolutional Transformer for No-Reference Video Quality AssessmentJunyong YouACM MM 2021 · 被引用 46 次
- Quality-aware Pretrained Models for Blind Image Quality AssessmentKai Zhao, Kun Yuan, Ming Sun, Mading Li 等CVPR 2023
- MetaIQA: Deep Meta-Learning for No-Reference Image Quality AssessmentHancheng Zhu, Leida Li, Jinjian Wu, Weisheng Dong 等CVPR 2020
- DSL-FIQA: Assessing Facial Image Quality via Dual-Set Degradation Learning and Landmark-Guided TransformerWei-Ting Chen, Gurunandan Krishnan, Qiang Gao, Sy-Yen Kuo 等CVPR 2024
