rPPG-VQA: A Video Quality Assessment Framework for Unsupervised rPPG Training
Tianyang Dai, Ming Chang, Yan Chen, Yang Hu
Abstract
Unsupervised remote photoplethysmography (rPPG) promises to leverage unlabeled video data, but its potential is hindered by a critical challenge: training on low-quality "in-the-wild" videos severely degrades model performance. An essential step missing here is to assess the suitability of the videos for rPPG model learning before using them for the task. Existing video quality assessment (VQA) methods are mainly designed for human perception and not directly applicable to the above purpose. In this work, we propose rPPG-VQA, a novel framework for assessing video suitability for rPPG. We integrate signal-level and scene-level analyses and design a dual-branch assessment architecture. The signal-level branch evaluates the physiological signal quality of the videos via robust signal-to-noise ratio (SNR) estimation with a multi-method consensus mechanism, and the scene-level branch uses a multimodal large language model (MLLM) to identify interferences like motion and unstable lighting. Furthermore, we propose a two-stage adaptive sampling (TAS) strategy that utilizes the quality score to curate optimal training datasets. Experiments show that by training on large-scale, "in-the-wild" videos filtered by our framework, we can develop unsupervised rPPG models that achieve a substantial improvement in accuracy on standard benchmarks. Our code is available at https://github.com/Tianyang-Dai/rPPG- VQA.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext a5fa06b3-5670-4035-9926-5542c470300eBuilds on8
- Exploring Video Quality Assessment on User Generated Contents from Aesthetic and Technical PerspectivesHaoning Wu, Erli Zhang, Liang Liao, Chaofeng Chen et al.ICCV 2023 · 371 citations
- Remote Heart Rate Measurement From Highly Compressed Facial Videos: An End-to-End Deep Learning Solution With Video EnhancementZitong Yu, Wei Peng, Xiaobai Li, Xiaopeng Hong et al.ICCV 2019 · 324 citations
- The Way to my Heart is through Contrastive Learning: Remote Photoplethysmography from Unlabelled VideoJohn Gideon, Simon StentICCV 2021 · 153 citations
- Modular Blind Video Quality AssessmentWen Wen, Mu Li, Yabin Zhang, Yiting Liao et al.CVPR 2024 · 23 citations
- G-Refine: A General Quality Refiner for Text-to-Image GenerationChunyi Li, Haoning Wu, Hongkun Hao, Zicheng Zhang et al.ACM MM 2024 · 7 citations
Related papers
- Non-Contrastive Unsupervised Learning of Physiological Signals from VideoJeremy Speth, Nathan Vance, Patrick J. Flynn, Adam CzajkaCVPR 2023
- Contactless Pulse Estimation Leveraging Pseudo Labels and Self-SupervisionZhihua Li, Lijun YinICCV 2023 · 21 citations
- PhysLLM: Harnessing Large Language Models for Cross-Modal Remote Physiological SensingYiping Xie, Bo Zhao, Mingtong Dai, Jian-Ping Zhou et al.ICLR 2026 · 19 citations
- Remote Photoplethysmography in Real-World and Extreme Lighting ScenariosHang Shao, Lei Luo, Jianjun Qian, Mengkai Yan et al.CVPR 2025
- Synthetic Generation of Face Videos with Plethysmograph PhysiologyZhen Wang, Yunhao Ba, Pradyumna Chari, Oyku Deniz Bozkurt et al.CVPR 2022 · 37 citations
