RL-ScanIQA: Reinforcement-Learned Scanpaths for Blind 360deg Image Quality Assessment
Yujia Wang, Yuyan Li, Jiuming Liu, Fang-Lue Zhang, Xinhu Zheng, Neil. A Dodgson
Abstract
Blind 360° image quality assessment (IQA) aims to predict perceptual quality for panoramic images without a pristine reference. Unlike conventional planar images, 360° content in immersive environments restricts viewers to a limited viewport at any moment, making viewing behaviors critical to quality perception. Although existing scanpath-based approaches have attempted to model viewing behaviors by approximating the human view‑then‑rate paradigm, they treat scanpath generation and quality assessment as separate steps, preventing end-to-end optimization and task-aligned exploration. To address this limitation, we propose RL‑ScanIQA, a reinforcement‑learned framework for blind 360° IQA. RL-ScanIQA optimize a PPO-trained scanpath policy and a quality assessor, where the policy receives quality-driven feedback to learn task-relevant viewing strategies. To improve training stability and prevent mode collapse, we design multi-level rewards, including scanpath diversity and equator-biased priors. We further boost cross‑dataset robustness using distortion‑space augmentation together with rank‑consistent losses that preserve intra‑image and inter‑image quality orderings. Extensive experiments on three benchmarks show that RL‑ScanIQA achieves superior in‑dataset performance and cross‑dataset generalization. Code will be released upon publication.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Builds on22
- MUSIQ: Multi-scale Image Quality TransformerJunjie Ke, Qifei Wang, Yilin Wang, Peyman Milanfar et al.ICCV 2021 · 1,325 citations
- Exploring CLIP for Assessing the Look and Feel of ImagesJianyi Wang, Kelvin C. K. Chan, Chen Change LoyAAAI 2023 · 1,208 citations
- Learning to summarize with human feedbackNisan Stiennon, Long Ouyang, Jeffrey Wu, Daniel M. Ziegler et al.NeurIPS 2020 · 124 citations
- Q-Insight: Understanding Image Quality via Visual Reinforcement LearningWeiqi Li, Xuanyu Zhang, Shijie Zhao, Yabin Zhang et al.NeurIPS 2025 · 117 citations
- RegFormer: An Efficient Projection-Aware Transformer Network for Large-Scale Point Cloud RegistrationJiuming Liu, Guangming Wang, Zhe Liu, Chaokang Jiang et al.ICCV 2023 · 71 citations
Related papers
- Learned Scanpaths Aid Blind Panoramic Video Quality AssessmentKanglong Fan, Wen Wen, Mu Li, Yifan Peng et al.CVPR 2024
- Assessor360: Multi-sequence Network for Blind Omnidirectional Image Quality AssessmentTianhe Wu, Shuwei Shi, Haoming Cai, Mingdeng Cao et al.NeurIPS 2023 · 57 citations
- TVFormer: Trajectory-guided Visual Quality Assessment on 360° Images with TransformersLi Yang, Mai Xu, Tie Liu, Liangyu Huo et al.ACM MM 2022 · 22 citations
- PanoEnv: Exploring 3D Spatial Intelligence in Panoramic Environments with Reinforcement LearningZekai Lin, Xu ZhengCVPR 2026 · 7 citations
- R3-PCQA: Ray-Reprojection-Reinforcement for No-Reference 3D Point Cloud Quality AssessmentJunhyuk Seo, Sanghyuk SEO, Dawoon Kim, Heeseok OhCVPR 2026
