CardioLive: Empowering Video Streaming with Online Cardiac Monitoring via Audio-Visual Learning
Sheng Lyu, Ruiming Huang, Sijie Ji, Yasar Abbas Ur Rehman, Lan Ma, Chenshu Wu
摘要
Online Cardiac Monitoring (OCM) emerges as a compelling enhancement for the next-generation video streaming platforms. It enables various applications, including remote health, affective computing, and deepfake detection. Yet the physiological information encapsulated in the video streams has long been neglected. In this paper, we present the design and implementation of CardioLive, the first online cardiac monitoring system in video streaming platforms. We leverage the naturally co-existing video and audio streams and devise CardioNet, the first audio-visual network to learn the cardiac series. It incorporates multiple unique designs to extract temporal and spectral features, ensuring robust performance under realistic streaming conditions. To enable the Service-On-Demand OCM, we implement CardioLive as a plug-and-play middleware service and develop systematic solutions to practical issues including changing FPS and unsynchronized streams. Extensive evaluations demonstrate the effectiveness of our system. We achieve a Mean Squared Error of 1.79 BPM error, outperforming the videoonly and audio-only solutions by 69.2% and 81.2%, respectively. CardioLive achieves average throughput of 115.97 and 98.16 FPS in Zoom and YouTube. We believe our work opens up new applications for video stream systems. Code is available at https://github.com/aiot-lab/CardioLive.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper40
- VideoMAE: Masked Autoencoders are Data-Efficient Learners for Self-Supervised Video Pre-TrainingZhan Tong, Yibing Song, Jue Wang, Limin WangNeurIPS 2022 · 被引用 2,336 次
- Multi-Task Temporal Shift Attention Networks for On-Device Contactless Vitals MeasurementXin Liu, Josh Fromm, Shwetak N. Patel, Daniel McDuffNeurIPS 2020 · 被引用 436 次
- Remote Heart Rate Measurement From Highly Compressed Facial Videos: An End-to-End Deep Learning Solution With Video EnhancementZitong Yu, Wei Peng, Xiaobai Li, Xiaopeng Hong 等ICCV 2019 · 被引用 324 次
- Server-Driven Video Streaming for Deep Learning InferenceKuntai Du, Ahsan Pervaiz, Xin Yuan, Aakanksha Chowdhery 等SIGCOMM 2020 · 被引用 238 次
- DeepRhythm: Exposing DeepFakes with Attentional Visual Heartbeat RhythmsHua Qi, Qing Guo, Felix Juefei-Xu, Xiaofei Xie 等ACM MM 2020 · 被引用 224 次
相关 Paper
- Emotions Don't Lie: An Audio-Visual Deepfake Detection Method using Affective CuesTrisha Mittal, Uttaran Bhattacharya, Rohan Chandra, Aniket Bera 等ACM MM 2020 · 被引用 314 次
- Stop My Dancing! Understanding, Detecting and Attributing Motion-Aware Deepfake VideosFazhong Liu, Yan Meng, Tian Dong, Guoxing Chen 等CCS 2026
- LiveScreen: Video Chat Liveness Detection Leveraging Skin ReflectionHongbo Liu, Zhihua Li, Yucheng Xie, Ruizhe Jiang 等INFOCOM 2020 · 被引用 13 次
- AITransfer: Progressive AI-powered Transmission for Real-Time Point Cloud Video StreamingYakun Huang, Yuanwei Zhu, Xiuquan Qiao, Zhijie Tan 等ACM MM 2021 · 被引用 34 次
- EduLive: Re-Creating Cues for Instructor-Learners Interaction in Educational Live Streams with Learners' Transcript-Based AnnotationsJingchao Fang, Jeongeon Park, Juho Kim, Hao-Chuan WangCSCW 2024 · 被引用 1 次
