Automated Class Discovery and One-Shot Interactions for Acoustic Activity Recognition
Jason Wu, Chris Harrison, Jeffrey P. Bigham, Gierad Laput
Abstract
Acoustic activity recognition has emerged as a foundational element for imbuing devices with context-driven capabilities, enabling richer, more assistive, and more accommodating computational experiences. Traditional approaches rely either on custom models trained in situ, or general models pre-trained on preexisting data, with each approach having accuracy and user burden implications. We present Listen Learner, a technique for activity recognition that gradually learns events specific to a deployed environment while minimizing user burden. Specifically, we built an end-to-end system for self-supervised learning of events labelled through one-shot interaction. We describe and quantify system performance 1) on preexisting audio datasets, 2) on real-world datasets we collected, and 3) through user studies which uncovered system behaviors suitable for this new type of interaction. Our results show that our system can accurately and automatically learn acoustic events across environments (e.g., 97% precision, 87% recall), while adhering to users' preferences for non-intrusive interactive behavior.
Ask about this paper
Ask your agent about it.
Lune has read the top-tier papers around this one, so every answer names the papers it rests on.
Your agent calls
Lunesearch_papers
Free to start. No credit card required.
Terminal
Install the CLIlune papers get 7b0bf42b-7316-42d2-a48b-6474652265ffCited by top-tier papers12
- Vid2Doppler: Synthesizing Doppler Radar Data from Videos for Training Privacy-Preserving Activity RecognitionKaran Ahuja, Yue Jiang, Mayank Goel, Chris HarrisonCHI 2021 · 118 citations
- Enabling Hand Gesture Customization on Wrist-Worn DevicesXuhai Xu, Jun Gong, Carolina Brum, Lilian Liang et al.CHI 2022 · 82 citations
- Exploring Spatial UI Transition Mechanisms with Head-Worn Augmented RealityFeiyu Lu, Yan XuCHI 2022 · 52 citations
- ProtoSound: A Personalized and Scalable Sound Recognition System for Deaf and Hard-of-Hearing UsersDhruv Jain, Khoa Huynh Anh Nguyen, Steven M. Goodman, Rachel Grossman-Kahn et al.CHI 2022 · 45 citations
- LipLearner: Customizable Silent Speech Interactions on Mobile DevicesZixiong Su, Shitao Fang, Jun RekimotoCHI 2023 · 37 citations
Related papers
- EchoLIFE: Zero-Shot In-Home ADL Recognition with LLM-Guided Active Acoustic SensingYubin Lan, Qian Zhang, Shukai Ma, Changfei Dong et al.UbiComp 2026
- Toward User-Driven Sound Recognizer Personalization with People Who Are d/Deaf or Hard of HearingSteven Goodman, Ping Liu, Dhruv Jain, Emma J. McDonnell et al.UbiComp 2021 · 33 citations
- EchoScriptor: Automatic Lifelogging Narratives via Activity-Based Audio-Language ModelKaylee Yaxuan Li, Xinghao Zhou, Haizhong Zheng, Kang G. Shin et al.CHI 2026
- Ok Google, What Am I Doing?: Acoustic Activity Recognition Bounded by Conversational Assistant InteractionsRebecca Adaimi, Howard Yong, Edison ThomazUbiComp 2021 · 29 citations
- ActivitySeeker: Towards Collaborative Personalized Human Activity Discovery and Recognition on SmartphonesZhoutong Ye, Yanwen Huang, Chun Yu, Yuntao Wang et al.CHI 2026 · 1 citation
