IMUTube: Automatic Extraction of Virtual on-body Accelerometry from Video for Human Activity Recognition
HyeokHyen Kwon, Catherine Tong, Harish Haresamudram, Yan Gao, Gregory D. Abowd, Nicholas D. Lane, Thomas Plötz
Abstract
The lack of large-scale, labeled data sets impedes progress in developing robust and generalized predictive models for on-body sensor-based human activity recognition (HAR). Labeled data in human activity recognition is scarce and hard to come by, as sensor data collection is expensive, and the annotation is time-consuming and error-prone. To address this problem, we introduce IMUTube, an automated processing pipeline that integrates existing computer vision and signal processing techniques to convert videos of human activity into virtual streams of IMU data. These virtual IMU streams represent accelerometry at a wide variety of locations on the human body. We show how the virtually-generated IMU data improves the performance of a variety of models on known HAR datasets. Our initial results are very promising, but the greater promise of this work lies in a collective approach by the computer vision, signal processing, and activity recognition communities to extend this work in ways that we outline. This should lead to on-body, sensor-based HAR becoming yet another success story in large-dataset breakthroughs in recognition.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext ac7d8d96-4177-4672-9e1c-18541253dfcdCited by top-tier papers27
- Contrastive Predictive Coding for Human Activity RecognitionHarish Haresamudram, Irfan A. Essa, Thomas PlötzUbiComp 2021 · 149 citations
- Vid2Doppler: Synthesizing Doppler Radar Data from Videos for Training Privacy-Preserving Activity RecognitionKaran Ahuja, Yue Jiang, Mayank Goel, Chris HarrisonCHI 2021 · 118 citations
- Assessing the State of Self-Supervised Human Activity Recognition Using WearablesHarish Haresamudram, Irfan Essa, Thomas PlötzUbiComp 2022 · 104 citations
- Towards Generalized mmWave-based Human Pose Estimation through Signal AugmentationHongfei Xue, Qiming Cao, Chenglin Miao, Yan Ju et al.MobiCom 2023 · 69 citations
- IMUGPT 2.0: Language-Based Cross Modality Transfer for Sensor-Based Human Activity RecognitionZikang Leng, Amitrajit Bhattacharjee, Hrudhai Rajasekhar, Lizhe Zhang et al.UbiComp 2024 · 59 citations
Builds on3
- AMASS: Archive of Motion Capture As Surface ShapesNaureen Mahmood, Nima Ghorbani, Nikolaus F. Troje, Gerard Pons-Moll et al.ICCV 2019 · 1,784 citations
- Depth From Videos in the Wild: Unsupervised Monocular Depth Learning From Unknown CamerasAriel Gordon, Hanhan Li, Rico Jonschkowski, Anelia AngelovaICCV 2019 · 397 citations
- Human-Aware Motion DeblurringZiyi Shen, Wenguan Wang, Xiankai Lu, Jianbing Shen et al.ICCV 2019 · 374 citations
Related papers
- Approaching the Real-World: Supporting Activity Recognition Training with Virtual IMU DataHyeokHyen Kwon, Bingyao Wang, Gregory D. Abowd, Thomas PlötzUbiComp 2021 · 48 citations
- Synthetic Smartwatch IMU Data Generation from In-the-wild ASL VideosPanneer Selvam Santhalingam, Parth Pathak, Huzefa Rangwala, Jana KoseckaUbiComp 2023 · 28 citations
- Practically Adopting Human Activity RecognitionHuatao Xu, Pengfei Zhou, Rui Tan, Mo LiMobiCom 2023 · 53 citations
- One Model to Fit Them All: Universal IMU-based Human Activity Recognition with LLM-assisted Cross-dataset RepresentationQingxin Wei, Jiaming Huang, Yi Gao, Wei DongUbiComp 2025 · 4 citations
- Vsens: Incorporating XR into the Process of Collecting Virtual IMU DataFengzhou Liang, Tian Min, Chengshuo Xia, Yuta SugiuraUbiComp 2026 · 1 citation
