IMUTube: Automatic Extraction of Virtual on-body Accelerometry from Video for Human Activity Recognition
HyeokHyen Kwon, Catherine Tong, Harish Haresamudram, Yan Gao, Gregory D. Abowd, Nicholas D. Lane, Thomas Plötz
摘要
The lack of large-scale, labeled data sets impedes progress in developing robust and generalized predictive models for on-body sensor-based human activity recognition (HAR). Labeled data in human activity recognition is scarce and hard to come by, as sensor data collection is expensive, and the annotation is time-consuming and error-prone. To address this problem, we introduce IMUTube, an automated processing pipeline that integrates existing computer vision and signal processing techniques to convert videos of human activity into virtual streams of IMU data. These virtual IMU streams represent accelerometry at a wide variety of locations on the human body. We show how the virtually-generated IMU data improves the performance of a variety of models on known HAR datasets. Our initial results are very promising, but the greater promise of this work lies in a collective approach by the computer vision, signal processing, and activity recognition communities to extend this work in ways that we outline. This should lead to on-body, sensor-based HAR becoming yet another success story in large-dataset breakthroughs in recognition.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper27
- Contrastive Predictive Coding for Human Activity RecognitionHarish Haresamudram, Irfan A. Essa, Thomas PlötzUbiComp 2021 · 被引用 149 次
- Vid2Doppler: Synthesizing Doppler Radar Data from Videos for Training Privacy-Preserving Activity RecognitionKaran Ahuja, Yue Jiang, Mayank Goel, Chris HarrisonCHI 2021 · 被引用 118 次
- Assessing the State of Self-Supervised Human Activity Recognition Using WearablesHarish Haresamudram, Irfan Essa, Thomas PlötzUbiComp 2022 · 被引用 104 次
- Towards Generalized mmWave-based Human Pose Estimation through Signal AugmentationHongfei Xue, Qiming Cao, Chenglin Miao, Yan Ju 等MobiCom 2023 · 被引用 69 次
- IMUGPT 2.0: Language-Based Cross Modality Transfer for Sensor-Based Human Activity RecognitionZikang Leng, Amitrajit Bhattacharjee, Hrudhai Rajasekhar, Lizhe Zhang 等UbiComp 2024 · 被引用 59 次
它引用的顶会 Paper3
- AMASS: Archive of Motion Capture As Surface ShapesNaureen Mahmood, Nima Ghorbani, Nikolaus F. Troje, Gerard Pons-Moll 等ICCV 2019 · 被引用 1,784 次
- Depth From Videos in the Wild: Unsupervised Monocular Depth Learning From Unknown CamerasAriel Gordon, Hanhan Li, Rico Jonschkowski, Anelia AngelovaICCV 2019 · 被引用 397 次
- Human-Aware Motion DeblurringZiyi Shen, Wenguan Wang, Xiankai Lu, Jianbing Shen 等ICCV 2019 · 被引用 374 次
相关 Paper
- Approaching the Real-World: Supporting Activity Recognition Training with Virtual IMU DataHyeokHyen Kwon, Bingyao Wang, Gregory D. Abowd, Thomas PlötzUbiComp 2021 · 被引用 48 次
- Synthetic Smartwatch IMU Data Generation from In-the-wild ASL VideosPanneer Selvam Santhalingam, Parth Pathak, Huzefa Rangwala, Jana KoseckaUbiComp 2023 · 被引用 28 次
- Practically Adopting Human Activity RecognitionHuatao Xu, Pengfei Zhou, Rui Tan, Mo LiMobiCom 2023 · 被引用 53 次
- One Model to Fit Them All: Universal IMU-based Human Activity Recognition with LLM-assisted Cross-dataset RepresentationQingxin Wei, Jiaming Huang, Yi Gao, Wei DongUbiComp 2025 · 被引用 4 次
- Vsens: Incorporating XR into the Process of Collecting Virtual IMU DataFengzhou Liang, Tian Min, Chengshuo Xia, Yuta SugiuraUbiComp 2026 · 被引用 1 次
