Synthetic Smartwatch IMU Data Generation from In-the-wild ASL Videos
Panneer Selvam Santhalingam, Parth Pathak, Huzefa Rangwala, Jana Kosecka
摘要
The scarcity of training data available for IMUs in wearables poses a serious challenge for IMU-based American Sign Language (ASL) recognition. In this paper, we ask the following question: can we "translate" the large number of publicly available, in-the-wild ASL videos to their corresponding IMU data? We answer this question by presenting a video to IMU translation framework (Vi2IMU) that takes as input user videos and estimates the IMU acceleration and gyro from the perspective of user's wrist. Vi2IMU consists of two modules, a wrist orientation estimation module that accounts for wrist rotations by carefully incorporating hand joint positions, and an acceleration and gyro prediction module, that leverages the orientation for transformation while capturing the contributions of hand movements and shape to produce realistic wrist acceleration and gyro data. We evaluate Vi2IMU by translating publicly available ASL videos to their corresponding wrist IMU data and train a gesture recognition model purely using the translated data. Our results show that the model using translated data performs reasonably well compared to the same model trained using measured IMU data.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- IMUGPT 2.0: Language-Based Cross Modality Transfer for Sensor-Based Human Activity RecognitionZikang Leng, Amitrajit Bhattacharjee, Hrudhai Rajasekhar, Lizhe Zhang 等UbiComp 2024 · 被引用 59 次
- AutoAugHAR: Automated Data Augmentation for Sensor-based Human Activity RecognitionYexu Zhou, Haibin Zhao, Yiran Huang, Tobias Röddiger 等UbiComp 2024 · 被引用 25 次
- SIGMA-ASL: Sensor-Integrated Multimodal Dataset for Sign Language RecognitionXiaofang Xiao, Guangchao Li, Guangrong Zhao, Qi Lin 等UbiComp 2026
它引用的顶会 Paper9
- IMUTube: Automatic Extraction of Virtual on-body Accelerometry from Video for Human Activity RecognitionHyeokHyen Kwon, Catherine Tong, Harish Haresamudram, Yan Gao 等UbiComp 2020 · 被引用 153 次
- Vid2Doppler: Synthesizing Doppler Radar Data from Videos for Training Privacy-Preserving Activity RecognitionKaran Ahuja, Yue Jiang, Mayank Goel, Chris HarrisonCHI 2021 · 被引用 118 次
- Approaching the Real-World: Supporting Activity Recognition Training with Virtual IMU DataHyeokHyen Kwon, Bingyao Wang, Gregory D. Abowd, Thomas PlötzUbiComp 2021 · 被引用 48 次
- IMU2Doppler: Cross-Modal Domain Adaptation for Doppler-based Activity Recognition Using IMU DataSejal Bhalla, Mayank Goel, Rushil KhuranaUbiComp 2022 · 被引用 42 次
- Teaching RF to Sense without RF Training MeasurementsHong Cai, Belal Korany, Chitra R. Karanam, Yasamin MostofiUbiComp 2021 · 被引用 42 次
相关 Paper
- SignRing: Continuous American Sign Language Recognition Using IMU Rings and Virtual IMU DataJiyang Li, Lin Huang, Siddharth Shah, Sean J. Jones 等UbiComp 2023 · 被引用 27 次
- WearSign: Pushing the Limit of Sign Language Translation Using Inertial and EMG WearablesQian Zhang, JiaZhen Jing, Dong Wang, Run ZhaoUbiComp 2022 · 被引用 26 次
- SmartASL: "Point-of-Care" Comprehensive ASL Interpreter Using WearablesYincheng Jin, Shibo Zhang, Yang Gao, Xuhai Xu 等UbiComp 2023 · 被引用 13 次
- iRadar: Synthesizing Millimeter-Waves from Wearable Inertial Inputs for Human Gesture SensingHuanqi Yang, Mingda Han, Xinyue Li, Di Duan 等INFOCOM 2025 · 被引用 9 次
- MultiHGR: Multi-Task Hand Gesture Recognition with Cross-Modal Wrist-Worn DevicesMengxia Lyu, Hao Zhou, Kaiwen Guo, Wangqiu Zhou 等INFOCOM 2024 · 被引用 4 次
