Self-Supervised Representation Learning Framework for Remote Physiological Measurement Using Spatiotemporal Augmentation Loss
Hao Wang, Euijoon Ahn, Jinman Kim
Abstract
Recent advances in supervised deep learning methods are enabling remote measurements of photoplethysmography-based physiological signals using facial videos. The performance of these supervised methods, however, are dependent on the availability of large labelled data. Contrastive learning as a self-supervised method has recently achieved state-of-the-art performances in learning representative data features by maximising mutual information between different augmented views. However, existing data augmentation techniques for contrastive learning are not designed to learn physiological signals from videos and often fail when there are complicated noise and subtle and periodic colour/shape variations between video frames. To address these problems, we present a novel self-supervised spatiotemporal learning framework for remote physiological signal representation learning, where there is a lack of labelled training data. Firstly, we propose a landmark-based spatial augmentation that splits the face into several informative parts based on the Shafer’s dichromatic reflection model to characterise subtle skin colour fluctuations. We also formulate a sparsity-based temporal augmentation exploiting Nyquist–Shannon sampling theorem to effectively capture periodic temporal changes by modelling physiological signal features. Furthermore, we introduce a constrained spatiotemporal loss which generates pseudo-labels for augmented video clips. It is used to regulate the training process and handle complicated noise. We evaluated our framework on 3 public datasets and demonstrated superior performances than other self-supervised methods and achieved competitive accuracy compared to the state-of-the-art supervised methods. Code is available at https://github.com/Dylan-H-Wang/SLF-RPM.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 22cef30e-b7f4-48ab-84a8-a2f4f0f462eeCited by top-tier papers5
- SimPer: Simple Self-Supervised Learning of Periodic TargetsYuzhe Yang, Xin Liu, Jiang Wu, Silviu Borac et al.ICLR 2023 · 20 citations
- An Unsupervised Approach for Periodic Source Detection in Time SeriesBerken Utku Demirel, Christian HolzICML 2024 · 3 citations
- Rhythmguassian: Repurposing Generalizable Gaussian Model for Remote Physiological MeasurementHao Lu, Yuting Zhang, Jiaqi Tang, Bowen Fu et al.ICCV 2025 · 2 citations
- Neuron Structure Modeling for Generalizable Remote Physiological MeasurementHao Lu, Zitong Yu, Xuesong Niu, Yingcong ChenCVPR 2023
- Non-Contrastive Unsupervised Learning of Physiological Signals from VideoJeremy Speth, Nathan Vance, Patrick J. Flynn, Adam CzajkaCVPR 2023
Builds on4
- Self-supervised Co-Training for Video Representation LearningTengda Han, Weidi Xie, Andrew ZissermanNeurIPS 2020 · 405 citations
- Remote Heart Rate Measurement From Highly Compressed Facial Videos: An End-to-End Deep Learning Solution With Video EnhancementZitong Yu, Wei Peng, Xiaobai Li, Xiaopeng Hong et al.ICCV 2019 · 324 citations
- Dual-GAN: Joint BVP and Noise Modeling for Remote Physiological MeasurementHao Lu, Hu Han, S. Kevin ZhouCVPR 2021
- Momentum Contrast for Unsupervised Visual Representation LearningKaiming He, Haoqi Fan, Yuxin Wu, Saining Xie et al.CVPR 2020
Related papers
- The Way to my Heart is through Contrastive Learning: Remote Photoplethysmography from Unlabelled VideoJohn Gideon, Simon StentICCV 2021 · 153 citations
- Contactless Pulse Estimation Leveraging Pseudo Labels and Self-SupervisionZhihua Li, Lijun YinICCV 2023 · 21 citations
- Spatiotemporal Contrastive Video Representation LearningRui Qian, Tianjian Meng, Boqing Gong, Ming-Hsuan Yang et al.CVPR 2021
- Cluster-Phys: Facial Clues Clustering Towards Efficient Remote Physiological MeasurementWei Qian, Kun Li, Dan Guo, Bin Hu et al.ACM MM 2024 · 19 citations
- Exploiting Self-Supervised and Semi-Supervised Learning for Facial Landmark Tracking with Unlabeled DataShi Yin, Shangfei Wang, Xiaoping Chen, Enhong ChenACM MM 2020 · 7 citations
