RIRNet: Recurrent-In-Recurrent Network for Video Quality Assessment
Pengfei Chen, Leida Li, Lei Ma, Jinjian Wu, Guangming Shi
Abstract
Video quality assessment (VQA), which is capable of automatically predicting the perceptual quality of source videos especially when reference information is not available, has become a major concern for video service providers due to the growing demand for video quality of experience (QoE) by end users. While significant advances have been achieved from the recent deep learning techniques, they often lead to misleading results in VQA tasks given their limitations on describing 3D spatio-temporal regularities using only fixed temporal frequency. Partially inspired by psychophysical and vision science studies revealing the speed tuning property of neurons in visual cortex when performing motion perception (i.e., sensitive to different temporal frequencies), we propose a novel no-reference (NR) VQA framework named Recurrent-In-Recurrent Network (RIRNet) to incorporate this characteristic to prompt an accurate representation of motion perception in VQA task. By fusing motion information derived from different temporal frequencies in a more efficient way, the resulting temporal modeling scheme is formulated to quantify the temporal motion effect via a hierarchical distortion description. It is found that the proposed framework is in closer agreement with quality perception of the distorted videos since it integrates concepts from motion perception in human visual system (HVS), which is manifested in the designed network structure composed of low- and high- level processing. A holistic validation of our methods on four challenging video quality databases demonstrates the superior performances over the state-of-the-art methods.
Ask about this paper
Ask your agent about it.
Lune has read the top-tier papers around this one, so every answer names the papers it rests on.
Your agent calls
Lunesearch_papers
Free to start. No credit card required.
Terminal
Install the CLIlune papers get f367e650-dc0d-4710-a4d1-d2711a448a3dCited by top-tier papers11
- Exploring Video Quality Assessment on User Generated Contents from Aesthetic and Technical PerspectivesHaoning Wu, Erli Zhang, Liang Liao, Chaofeng Chen et al.ICCV 2023 · 371 citations
- A Deep Learning based No-reference Quality Assessment Model for UGC VideosWei Sun, Xiongkuo Min, Wei Lu, Guangtao ZhaiACM MM 2022 · 239 citations
- Towards Explainable In-the-Wild Video Quality Assessment: A Database and a Language-Prompted ApproachHaoning Wu, Erli Zhang, Liang Liao, Chaofeng Chen et al.ACM MM 2023 · 51 citations
- Unsupervised Curriculum Domain Adaptation for No-Reference Video Quality AssessmentPengfei Chen, Leida Li, Jinjian Wu, Weisheng Dong et al.ICCV 2021 · 40 citations
- Exploring the Effectiveness of Video Perceptual Representation in Blind Video Quality AssessmentLiang Liao, Kangmin Xu, Haoning Wu, Chaofeng Chen et al.ACM MM 2022 · 39 citations
Related papers
- Long Short-term Convolutional Transformer for No-Reference Video Quality AssessmentJunyong YouACM MM 2021 · 46 citations
- Modular Blind Video Quality AssessmentWen Wen, Mu Li, Yabin Zhang, Yiting Liao et al.CVPR 2024 · 23 citations
- Semantic-Aware and Quality-Aware Interaction Network for Blind Video Quality AssessmentJianjun Xiang, Yuanjie Dang, Peng Chen, Ronghua Liang et al.ACM MM 2024
- Capturing Co-existing Distortions in User-Generated Content for No-reference Video Quality AssessmentKun Yuan, Zishang Kong, Chuanchuan Zheng, Ming Sun et al.ACM MM 2023 · 15 citations
- PTM-VQA: Efficient Video Quality Assessment Leveraging Diverse PreTrained Models from the WildKun Yuan, Hongbo Liu, Mading Li, Muyi Sun et al.CVPR 2024 · 8 citations
