Semantic-Aware and Quality-Aware Interaction Network for Blind Video Quality Assessment
Jianjun Xiang, Yuanjie Dang, Peng Chen, Ronghua Liang, Ruohong Huan, Nan Gao
摘要
Current state-of-the-art video quality assessment (VQA) models typically integrate various perceptual features to comprehensively represent video quality degradation. These models either directly concatenate features or fuse different perceptual scores while ignoring the domain gaps between cross-aware features, thus failing to adequately learn the correlations and interactions between different perceptual features. To this end, we analyze the independent effects and information gaps of quality-and semantic-aware features on video quality. Based on an analysis of the spatial and temporal differences between two aware features, we propose a semantic-Aware and quality-Aware Interaction Network (A2INet) for blind VQA. For spatial gaps, we introduce a cross-aware guided interaction module to enhance the interaction between semantic-and quality-aware features in a local-to-global manner. Considering temporal discrepancies, we design a cross-aware temporal modeling module to further perceive temporal content variation and quality saliency information, and perceptual features are regressed into quality score by a temporal network and a temporal pooling. Extensive experiments on six benchmark VQA datasets show that our model achieves state-of-the-art performance, and ablation studies further validate the effectiveness of each module. We also present a simple video sampling strategy to balance the effectiveness and efficiency of the model. The code for the proposed method will be released at https://github.com/JianjunXiang/A2INet.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
引用它的顶会 Paper2
- MVQA-68K: A Multi-dimensional and Causally-annotated Dataset with Quality Interpretability for Video AssessmentYanyun Pu, Kehan Li, Zeyi Huang, Zhijie Zhong 等ACM MM 2025 · 被引用 2 次
- DSP-PCQA: Integrating Multiple Perception Preferences for Point Cloud Quality AssessmentMingxuan Li, Fazhan Zhang, Zhenzhe Hou, Zihao Huang 等AAAI 2026
相关 Paper
- Modular Blind Video Quality AssessmentWen Wen, Mu Li, Yabin Zhang, Yiting Liao 等CVPR 2024 · 被引用 23 次
- A Deep Learning based No-reference Quality Assessment Model for UGC VideosWei Sun, Xiongkuo Min, Wei Lu, Guangtao ZhaiACM MM 2022 · 被引用 239 次
- Blind Natural Video Quality Prediction via Statistical Temporal Features and Deep Spatial FeaturesJari Korhonen, Yicheng Su, Junyong YouACM MM 2020 · 被引用 88 次
- ADGNet: Attention Discrepancy Guided Deep Neural Network for Blind Image Quality AssessmentXiaoyu Ma, Yaqi Wang, Chang Liu, Suiyu Zhang 等ACM MM 2022 · 被引用 6 次
- Multiview Contrastive Learning for Completely Blind Video Quality Assessment of User Generated ContentShankhanil Mitra, Rajiv SoundararajanACM MM 2022 · 被引用 9 次
