ProtoSound: A Personalized and Scalable Sound Recognition System for Deaf and Hard-of-Hearing Users
Dhruv Jain, Khoa Huynh Anh Nguyen, Steven M. Goodman, Rachel Grossman-Kahn, Hung Ngo, Aditya Kusupati, Ruofei Du, Alex Olwal, Leah Findlater, Jon E. Froehlich
摘要
Recent advances have enabled automatic sound recognition systems for deaf and hard of hearing (DHH) users on mobile devices. However, these tools use pre-trained, generic sound recognition models, which do not meet the diverse needs of DHH users. We introduce ProtoSound, an interactive system for customizing sound recognition models by recording a few examples, thereby enabling personalized and fine-grained categories. ProtoSound is motivated by prior work examining sound awareness needs of DHH people and by a survey we conducted with 472 DHH participants. To evaluate ProtoSound, we characterized performance on two real-world sound datasets, showing significant improvement over state-of-the-art (e.g., +9.7% accuracy on the first dataset). We then deployed ProtoSound's end-user training and real-time recognition through a mobile application and recruited 19 hearing participants who listened to the real-world sounds and rated the accuracy across 56 locations (e.g., homes, restaurants, parks). Results show that ProtoSound personalized the model on-device in real-time and accurately learned sounds across diverse acoustic contexts. We close by discussing open challenges in personalizable sound recognition, including the need for better recording interfaces and algorithmic improvements.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper8
- LipLearner: Customizable Silent Speech Interactions on Mobile DevicesZixiong Su, Shitao Fang, Jun RekimotoCHI 2023 · 被引用 37 次
- "Easier or Harder, Depending on Who the Hearing Person Is": Codesigning Videoconferencing Tools for Small Groups with Mixed Hearing StatusEmma J. McDonnell, Soo Hyun Moon, Lucy Jiang, Steven M. Goodman 等CHI 2023 · 被引用 26 次
- Exploring AI Problem Formulation with Children via Teachable MachinesUtkarsh Dwivedi, Salma Elsayed-Ali, Elizabeth Bonsignore, Hernisa KacorriCHI 2024 · 被引用 14 次
- SPECTRA: Personalizable Sound Recognition for Deaf and Hard of Hearing Users through Interactive Machine LearningSteven M. Goodman, Emma J. McDonnell, Jon E. Froehlich, Leah FindlaterCHI 2025 · 被引用 10 次
- Super Kawaii Vocalics: Amplifying the "Cute" Factor in Computer VoiceYuto Mandai, Katie Seaborn, Tomoyasu Nakano, Xin Sun 等CHI 2025 · 被引用 9 次
它引用的顶会 Paper5
- Rapid Learning or Feature Reuse? Towards Understanding the Effectiveness of MAMLAniruddh Raghu, Maithra Raghu, Samy Bengio, Oriol VinyalsICLR 2020 · 被引用 736 次
- Automated Class Discovery and One-Shot Interactions for Acoustic Activity RecognitionJason Wu, Chris Harrison, Jeffrey P. Bigham, Gierad LaputCHI 2020 · 被引用 52 次
- HomeSound: An Iterative Field Deployment of an In-Home Sound Awareness System for Deaf or Hard of Hearing UsersDhruv Jain, Kelly Mack, Akli Amrous, Matt Wright 等CHI 2020 · 被引用 52 次
- Evaluating Smartwatch-based Sound Feedback for Deaf and Hard-of-hearing Users Across ContextsSteven Goodman, Susanne Kirchner, Rose Guttman, Dhruv Jain 等CHI 2020 · 被引用 36 次
- Toward User-Driven Sound Recognizer Personalization with People Who Are d/Deaf or Hard of HearingSteven Goodman, Ping Liu, Dhruv Jain, Emma J. McDonnell 等UbiComp 2021 · 被引用 33 次
相关 Paper
- UbiHearo: Bringing Scenario-Aware Sound Guidance to DHH Users with Mobile AgentsFengmin Wu, Sicong Liu, Zimu Zhou, Chenren Xu 等UbiComp 2026
- A Human-AI Collaborative Approach for Designing Sound Awareness SystemsJeremy Zhengqi Huang, Reyna Wood, Hriday Chhabria, Dhruv JainCHI 2024 · 被引用 6 次
- It Didn't Sound Good with My Cochlear Implants: Understanding the Challenges of Using Smart Assistants for Deaf and Hard of Hearing UsersJohnna Blair, Saeed AbdullahUbiComp 2021 · 被引用 24 次
- SoundTrace: Integrating Temporal Context and Episodic Memory for Real-Time Sound RecognitionDhruv Jain, Jason MillerUbiComp 2026
- Analyzing Deaf and Hard-of-Hearing Users' Behavior, Usage, and Interaction with a Personal Assistant Device that Understands Sign-Language InputAbraham Glasser, Matthew Watkins, Kira Hart, Sooyeon Lee 等CHI 2022 · 被引用 21 次
