Self-Supervised Representation Learning Framework for Remote Physiological Measurement Using Spatiotemporal Augmentation Loss
Hao Wang, Euijoon Ahn, Jinman Kim
摘要
Recent advances in supervised deep learning methods are enabling remote measurements of photoplethysmography-based physiological signals using facial videos. The performance of these supervised methods, however, are dependent on the availability of large labelled data. Contrastive learning as a self-supervised method has recently achieved state-of-the-art performances in learning representative data features by maximising mutual information between different augmented views. However, existing data augmentation techniques for contrastive learning are not designed to learn physiological signals from videos and often fail when there are complicated noise and subtle and periodic colour/shape variations between video frames. To address these problems, we present a novel self-supervised spatiotemporal learning framework for remote physiological signal representation learning, where there is a lack of labelled training data. Firstly, we propose a landmark-based spatial augmentation that splits the face into several informative parts based on the Shafer’s dichromatic reflection model to characterise subtle skin colour fluctuations. We also formulate a sparsity-based temporal augmentation exploiting Nyquist–Shannon sampling theorem to effectively capture periodic temporal changes by modelling physiological signal features. Furthermore, we introduce a constrained spatiotemporal loss which generates pseudo-labels for augmented video clips. It is used to regulate the training process and handle complicated noise. We evaluated our framework on 3 public datasets and demonstrated superior performances than other self-supervised methods and achieved competitive accuracy compared to the state-of-the-art supervised methods. Code is available at https://github.com/Dylan-H-Wang/SLF-RPM.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- SimPer: Simple Self-Supervised Learning of Periodic TargetsYuzhe Yang, Xin Liu, Jiang Wu, Silviu Borac 等ICLR 2023 · 被引用 20 次
- An Unsupervised Approach for Periodic Source Detection in Time SeriesBerken Utku Demirel, Christian HolzICML 2024 · 被引用 3 次
- Rhythmguassian: Repurposing Generalizable Gaussian Model for Remote Physiological MeasurementHao Lu, Yuting Zhang, Jiaqi Tang, Bowen Fu 等ICCV 2025 · 被引用 2 次
- Neuron Structure Modeling for Generalizable Remote Physiological MeasurementHao Lu, Zitong Yu, Xuesong Niu, Yingcong ChenCVPR 2023
- Non-Contrastive Unsupervised Learning of Physiological Signals from VideoJeremy Speth, Nathan Vance, Patrick J. Flynn, Adam CzajkaCVPR 2023
它引用的顶会 Paper4
- Self-supervised Co-Training for Video Representation LearningTengda Han, Weidi Xie, Andrew ZissermanNeurIPS 2020 · 被引用 405 次
- Remote Heart Rate Measurement From Highly Compressed Facial Videos: An End-to-End Deep Learning Solution With Video EnhancementZitong Yu, Wei Peng, Xiaobai Li, Xiaopeng Hong 等ICCV 2019 · 被引用 324 次
- Dual-GAN: Joint BVP and Noise Modeling for Remote Physiological MeasurementHao Lu, Hu Han, S. Kevin ZhouCVPR 2021
- Momentum Contrast for Unsupervised Visual Representation LearningKaiming He, Haoqi Fan, Yuxin Wu, Saining Xie 等CVPR 2020
相关 Paper
- The Way to my Heart is through Contrastive Learning: Remote Photoplethysmography from Unlabelled VideoJohn Gideon, Simon StentICCV 2021 · 被引用 153 次
- Contactless Pulse Estimation Leveraging Pseudo Labels and Self-SupervisionZhihua Li, Lijun YinICCV 2023 · 被引用 21 次
- Spatiotemporal Contrastive Video Representation LearningRui Qian, Tianjian Meng, Boqing Gong, Ming-Hsuan Yang 等CVPR 2021
- Cluster-Phys: Facial Clues Clustering Towards Efficient Remote Physiological MeasurementWei Qian, Kun Li, Dan Guo, Bin Hu 等ACM MM 2024 · 被引用 19 次
- Exploiting Self-Supervised and Semi-Supervised Learning for Facial Landmark Tracking with Unlabeled DataShi Yin, Shangfei Wang, Xiaoping Chen, Enhong ChenACM MM 2020 · 被引用 7 次
