ColorVideoVDP: A visual difference predictor for image, video and display distortions
Rafal K. Mantiuk, Param Hanji, Maliha Ashraf, Yuta Asano, Alexandre Chapiro
Abstract
ColorVideoVDP is a video and image quality metric that models spatial and temporal aspects of vision for both luminance and color. The metric is built on novel psychophysical models of chromatic spatiotemporal contrast sensitivity and cross-channel contrast masking. It accounts for the viewing conditions, geometric, and photometric characteristics of the display. It was trained to predict common video-streaming distortions (e.g., video compression, rescaling, and transmission errors) and also 8 new distortion types related to AR/VR displays (e.g., light source and waveguide non-uniformities). To address the latter application, we collected our novel XR-Display-Artifact-Video quality dataset (XR-DAVID), comprised of 336 distorted videos. Extensive testing on XR-DAVID, as well as several datasets from the literature, indicate a significant gain in prediction performance compared to existing metrics. ColorVideoVDP opens the doors to many novel applications that require the joint automated spatiotemporal assessment of luminance and color distortions, including video streaming, display specification, and design, visual comparison of results, and perceptually-guided quality optimization. The code for the metric can be found at https://github.com/gfxdisp/ColorVideoVDP.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext ca6e8992-2f3c-4aa8-af5d-b91c300af171Cited by top-tier papers9
- What is HDR? Perceptual Impact of Luminance and Contrast in Immersive DisplaysKenneth Chen, Nathan Matsuda, Jon McElvain, Yang Zhao et al.SIGGRAPH 2025 · 4 citations
- Puzzle Similarity: A Perceptually-Guided Cross-Reference Metric for Artifact Detection in 3D Scene ReconstructionsNicolai Hermann, Jorge Condor, Piotr DidykICCV 2025 · 4 citations
- Velox: Learning Representations of 4D Geometry and AppearanceAnagh Malik, Dorian Chan, Xiaoming Zhao, David B. Lindell et al.CVPR 2026
- Streaming of rendered content with adaptive frame rate and resolutionYaru Liu, Joseph G. March, Rafal K. MantiukSIGGRAPH 2026
- Forget Superresolution, Sample Adaptively (when Path Tracing)Martin Bálint, Corentin Salaün, Hans-Peter Seidel, Karol MyszkowskiSIGGRAPH 2026
Builds on3
- FovVideoVDP: a visible difference predictor for wide field-of-view videoRafal K. Mantiuk, Gyorgy Denes, Alexandre Chapiro, Anton Kaplanyan et al.SIGGRAPH 2021 · 158 citations
- stelaCSF: a unified model of contrast sensitivity as the function of spatio-temporal frequency, eccentricity, luminance and areaRafal K. Mantiuk, Maliha Ashraf, Alexandre ChapiroSIGGRAPH 2022 · 56 citations
- A perceptual model of motion quality for rendering with adaptive refresh-rate and resolutionGyorgy Denes, Akshay Jindal, Aliaksei Mikhailiuk, Rafal K. MantiukSIGGRAPH 2020 · 40 citations
Related papers
- Adapting Quality Metrics to Tone MappingKenneth Chen, Dongyeon Kim, Yuta Asano, Alexandre Chapiro et al.SIGGRAPH 2026
- Patch-VQ: 'Patching Up' the Video Quality ProblemZhenqiang Ying, Maniratnam Mandal, Deepti Ghadiyaram, Alan C. BovikCVPR 2021
- Learning Flexible Generalization in Video Quality Assessment by Bringing Device and Viewing Condition DistributionsNickolay Safonov, Dmitriy VatolinICML 2026
- Inverse-Tone-Mapped HDR Video Quality Assessment for Broadcast Television: A Comprehensive Dataset and SDR-Referenced MethodLeidong Fan, Qian Zhang, Qing LiACM MM 2025
- Towards Explainable In-the-Wild Video Quality Assessment: A Database and a Language-Prompted ApproachHaoning Wu, Erli Zhang, Liang Liao, Chaofeng Chen et al.ACM MM 2023 · 51 citations
