Automated Class Discovery and One-Shot Interactions for Acoustic Activity Recognition
Jason Wu, Chris Harrison, Jeffrey P. Bigham, Gierad Laput
摘要
Acoustic activity recognition has emerged as a foundational element for imbuing devices with context-driven capabilities, enabling richer, more assistive, and more accommodating computational experiences. Traditional approaches rely either on custom models trained in situ, or general models pre-trained on preexisting data, with each approach having accuracy and user burden implications. We present Listen Learner, a technique for activity recognition that gradually learns events specific to a deployed environment while minimizing user burden. Specifically, we built an end-to-end system for self-supervised learning of events labelled through one-shot interaction. We describe and quantify system performance 1) on preexisting audio datasets, 2) on real-world datasets we collected, and 3) through user studies which uncovered system behaviors suitable for this new type of interaction. Our results show that our system can accurately and automatically learn acoustic events across environments (e.g., 97% precision, 87% recall), while adhering to users' preferences for non-intrusive interactive behavior.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
引用它的顶会 Paper12
- Vid2Doppler: Synthesizing Doppler Radar Data from Videos for Training Privacy-Preserving Activity RecognitionKaran Ahuja, Yue Jiang, Mayank Goel, Chris HarrisonCHI 2021 · 被引用 118 次
- Enabling Hand Gesture Customization on Wrist-Worn DevicesXuhai Xu, Jun Gong, Carolina Brum, Lilian Liang 等CHI 2022 · 被引用 82 次
- Exploring Spatial UI Transition Mechanisms with Head-Worn Augmented RealityFeiyu Lu, Yan XuCHI 2022 · 被引用 52 次
- ProtoSound: A Personalized and Scalable Sound Recognition System for Deaf and Hard-of-Hearing UsersDhruv Jain, Khoa Huynh Anh Nguyen, Steven M. Goodman, Rachel Grossman-Kahn 等CHI 2022 · 被引用 45 次
- LipLearner: Customizable Silent Speech Interactions on Mobile DevicesZixiong Su, Shitao Fang, Jun RekimotoCHI 2023 · 被引用 37 次
相关 Paper
- EchoLIFE: Zero-Shot In-Home ADL Recognition with LLM-Guided Active Acoustic SensingYubin Lan, Qian Zhang, Shukai Ma, Changfei Dong 等UbiComp 2026
- Toward User-Driven Sound Recognizer Personalization with People Who Are d/Deaf or Hard of HearingSteven Goodman, Ping Liu, Dhruv Jain, Emma J. McDonnell 等UbiComp 2021 · 被引用 33 次
- EchoScriptor: Automatic Lifelogging Narratives via Activity-Based Audio-Language ModelKaylee Yaxuan Li, Xinghao Zhou, Haizhong Zheng, Kang G. Shin 等CHI 2026
- Ok Google, What Am I Doing?: Acoustic Activity Recognition Bounded by Conversational Assistant InteractionsRebecca Adaimi, Howard Yong, Edison ThomazUbiComp 2021 · 被引用 29 次
- ActivitySeeker: Towards Collaborative Personalized Human Activity Discovery and Recognition on SmartphonesZhoutong Ye, Yanwen Huang, Chun Yu, Yuntao Wang 等CHI 2026 · 被引用 1 次
