SoundTrace: Integrating Temporal Context and Episodic Memory for Real-Time Sound Recognition
Dhruv Jain, Jason Miller
Abstract
Environmental sound recognition systems have become increasingly capable, yet they often operate in context-free modes that ignore temporal continuity, environmental patterns, and user feedback. We present SoundTrace, a real-time sound recognition system that integrates temporal context and episodic memory to support adaptive, interpretable inference in everyday environments. SoundTrace augments a neural audio classifier with lightweight memory structures that store symbolic event traces, estimate scene-time and short-range sequence priors, and update those priors through user feedback. During inference, these memory-derived priors are retrieved and fused with model predictions to stabilize labels and provide interpretable reasoning. In a controlled evaluation, we show that contextual inference improves accuracy, reduces label volatility, and enhances robustness under ambiguous conditions. We also report findings from an eight-week in-home deployment with 14 deaf and hard-of-hearing participants, revealing how context-aware feedback and explanations shape users' trust, understanding, and correction strategies. Our results demonstrate the viability of context-integrated sound recognition in everyday environments.
Ask about this paper
Ask your agent about it.
Lune has read the top-tier papers around this one, so every answer names the papers it rests on.
Related papers
- UbiHearo: Bringing Scenario-Aware Sound Guidance to DHH Users with Mobile AgentsFengmin Wu, Sicong Liu, Zimu Zhou, Chenren Xu et al.UbiComp 2026
- EchoScriptor: Automatic Lifelogging Narratives via Activity-Based Audio-Language ModelKaylee Yaxuan Li, Xinghao Zhou, Haizhong Zheng, Kang G. Shin et al.CHI 2026
- ProtoSound: A Personalized and Scalable Sound Recognition System for Deaf and Hard-of-Hearing UsersDhruv Jain, Khoa Huynh Anh Nguyen, Steven M. Goodman, Rachel Grossman-Kahn et al.CHI 2022 · 45 citations
- Automated Class Discovery and One-Shot Interactions for Acoustic Activity RecognitionJason Wu, Chris Harrison, Jeffrey P. Bigham, Gierad LaputCHI 2020 · 52 citations
- HomeSound: An Iterative Field Deployment of an In-Home Sound Awareness System for Deaf or Hard of Hearing UsersDhruv Jain, Kelly Mack, Akli Amrous, Matt Wright et al.CHI 2020 · 52 citations
