The EarSAVAS Dataset: Enabling Subject-Aware Vocal Activity Sensing on Earables
Xiyuxing Zhang, Yuntao Wang, Yuxuan Han, Chen Liang, Ishan Chatterjee, Jiankai Tang, Xin Yi, Shwetak N. Patel, Yuanchun Shi
Abstract
Subject-aware vocal activity sensing on wearables, which specifically recognizes and monitors the wearer's distinct vocal activities, is essential in advancing personal health monitoring and enabling context-aware applications. While recent advancements in earables present new opportunities, the absence of relevant datasets and effective methods remains a significant challenge. In this paper, we introduce EarSAVAS, the first publicly available dataset constructed specifically for subject-aware human vocal activity sensing on earables. EarSAVAS encompasses eight distinct vocal activities from both the earphone wearer and bystanders, including synchronous two-channel audio and motion data collected from 42 participants totaling 44.5 hours. Further, we propose EarVAS, a lightweight multi-modal deep learning architecture that enables efficient subject-aware vocal activity recognition on earables. To validate the reliability of EarSAVAS and the efficiency of EarVAS, we implemented two advanced benchmark models. Evaluation results on EarSAVAS reveal EarVAS's effectiveness with an accuracy of 90.84% and a Macro-AUC of 89.03%. Comprehensive ablation experiments were conducted on benchmark models and demonstrated the effectiveness of feedback microphone audio and highlighted the potential value of sensor fusion in subject-aware vocal activity sensing on earables. We hope that the proposed EarSAVAS and benchmark models can inspire other researchers to further explore efficient subject-aware human vocal activity sensing on earables.
Ask about this paper
Ask your agent about it.
Lune has read the top-tier papers around this one, so every answer names the papers it rests on.
Cited by top-tier papers4
- Vision-Based Multimodal Interfaces: A Survey and Taxonomy for Enhanced Context-Aware System DesignYongquan 'Owen' Hu, Jingyu Tang, Xinya Gong, Zhongyi Zhou et al.CHI 2025 · 37 citations
- A Survey of Earable Technology: Trends, Tools, and the Road AheadChangshuo Hu, Qiang Yang, Yang Liu, Tobias Röddiger et al.UbiComp 2026 · 4 citations
- SonicSieve: Bringing Directional Speech Extraction to Smartphones Using Acoustic MicrostructuresKuang Yuan, Yifeng Wang, Xiyuxing Zhang, Chengyi Shen et al.CHI 2026 · 1 citation
- Earinter: A Closed-Loop System for Eating Pace Regulation with Just-in-Time Intervention Using Commodity EarbudsJun Fang, Ka I. Chan, Xiyuxing Zhang, Yuntao Wang et al.UbiComp 2026
Related papers
- EarSleep: In-ear Acoustic-based Physical and Physiological Activity Recognition for Sleep Stage DetectionFeiyu Han, Panlong Yang, Yuanhao Feng, Weiwei Jiang et al.UbiComp 2024 · 16 citations
- Leveraging Sound and Wrist Motion to Detect Activities of Daily Living with Commodity SmartwatchesSarnab Bhattacharya, Rebecca Adaimi, Edison ThomazUbiComp 2022 · 41 citations
- Hearing Your Calories: Short-Session Calibrated Energy Expenditure Monitoring via Earable Respiratory SensingYetong Cao, Xiaochen Liu, Jianquan Zhao, Dong Ma et al.UbiComp 2026
- EarBuddy: Enabling On-Face Interaction via Wireless EarbudsXuhai Xu, Haitian Shi, Xin Yi, Wenjia Liu et al.CHI 2020 · 91 citations
- SAMoSA: Sensing Activities with Motion and Subsampled AudioVimal Mollyn, Karan Ahuja, Dhruv Verma, Chris Harrison et al.UbiComp 2022 · 54 citations
