SAMoSA: Sensing Activities with Motion and Subsampled Audio
Vimal Mollyn, Karan Ahuja, Dhruv Verma, Chris Harrison, Mayank Goel
摘要
Despite advances in audio- and motion-based human activity recognition (HAR) systems, a practical, power-efficient, and privacy-sensitive activity recognition system has remained elusive. State-of-the-art activity recognition systems often require power-hungry and privacy-invasive audio data. This is especially challenging for resource-constrained wearables, such as smartwatches. To counter the need for an always-on audio-based activity classification system, we first make use of power and compute-optimized IMUs sampled at 50 Hz to act as a trigger for detecting activity events. Once detected, we use a multimodal deep learning model that augments the motion data with audio data captured on a smartwatch. We subsample this audio to rates ≤ 1 kHz, rendering spoken content unintelligible, while also reducing power consumption on mobile devices. Our multimodal deep learning model achieves a recognition accuracy of 92.2% across 26 daily activities in four indoor environments. Our findings show that subsampling audio from 16 kHz down to 1 kHz, in concert with motion data, does not result in a significant drop in inference accuracy. We also analyze the speech content intelligibility and power requirements of audio sampled at less than 1 kHz and demonstrate that our proposed approach can improve the practicality of human activity recognition systems.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
引用它的顶会 Paper15
- EchoWrist: Continuous Hand Pose Tracking and Hand-Object Interaction Recognition Using Low-Power Active Acoustic Sensing On a WristbandChi-Jung Lee, Ruidong Zhang, Devansh Agarwal, Tianhong Catherine Yu 等CHI 2024 · 被引用 48 次
- Semantic Hearing: Programming Acoustic Scenes with Binaural HearablesBandhav Veluri, Malek Itani, Justin Chan, Takuya Yoshioka 等UIST 2023 · 被引用 29 次
- ActSonic: Recognizing Everyday Activities from Inaudible Acoustic Wave Around the BodySaif Mahmud, Vineet Parikh, Qikang Liang, Ke Li 等UbiComp 2025 · 被引用 24 次
- PrISM-Tracker: A Framework for Multimodal Procedure Tracking Using Wearable Sensors and State Transition Information with User-Driven Handling of Errors and UncertaintyRiku Arakawa, Hiromu Yakura, Vimal Mollyn, Suzanne Nie 等UbiComp 2023 · 被引用 20 次
- PrISM-Q&A: Step-Aware Voice Assistant on a Smartwatch Enabled by Multimodal Procedure Tracking and Large Language ModelsRiku Arakawa, Jill Fain Lehman, Mayank GoelUbiComp 2025 · 被引用 20 次
相关 Paper
- Leveraging Sound and Wrist Motion to Detect Activities of Daily Living with Commodity SmartwatchesSarnab Bhattacharya, Rebecca Adaimi, Edison ThomazUbiComp 2022 · 被引用 41 次
- IMU2Doppler: Cross-Modal Domain Adaptation for Doppler-based Activity Recognition Using IMU DataSejal Bhalla, Mayank Goel, Rushil KhuranaUbiComp 2022 · 被引用 42 次
- DeepSenseMoE: Harnessing Power of Time Series Foundation Models for Few-Shot Human Activity RecognitionZenan Fu, Dongzhou Cheng, Lei Zhang, Wenbo Huang 等AAAI 2026
- Unsupervised Human Activity Representation Learning with Multi-task Deep ClusteringHaojie Ma, Zhijie Zhang, Wenzhong Li, Sanglu LuUbiComp 2021 · 被引用 46 次
- Kirigami: Lightweight Speech Filtering for Privacy-Preserving Activity Recognition using AudioSudershan Boovaraghavan, Haozhe Zhou, Mayank Goel, Yuvraj AgarwalUbiComp 2024 · 被引用 7 次
