Digital Ventriloquism: Giving Voice to Everyday Objects
Yasha Iravantchi, Mayank Goel, Chris Harrison
摘要
Smart speakers with voice agents are becoming increasingly common. However, the agent's voice always emanates from the device, even when that information is contextually and spatially relevant elsewhere. Digital Ventriloquism allows smart speakers to render sound onto everyday objects, such that it appears they are speaking and are interactive. This can be achieved without any modification of objects or the environment. For this, we used a highly directional pan-tilt ultrasonic array. By modulating a 40 kHz ultrasonic signal, we can emit sound that is inaudible "in flight" and demodulates to audible frequencies when impacting a surface through acoustic parametric interaction. This makes it appear as though the sound originates from an object and not the speaker. We ran a study in which we projected speech onto five objects in three environments, and found that participants were able to correctly identify the source object 92% of the time and correctly repeat the spoken message 100% of the time, demonstrating our digital ventriloquy is both directional and intelligible.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper6
- Evoking empathy with visually impaired people through an augmented reality embodiment experienceRenan Guarese, Emma Pretty, Haytham M. Fayek, Fabio Zambetta 等IEEE VR 2023 · 被引用 23 次
- Auptimize: Optimal Placement of Spatial Audio Cues for Extended RealityHyunsung Cho, Alexander Wang, Divya Kartik, Emily Liying Xie 等UIST 2024 · 被引用 19 次
- Bubbleu: Exploring Augmented Reality Game Design with Uncertain AI-based InteractionMinji Kim, Kyungjin Lee, Rajesh Balan, Youngki LeeCHI 2023 · 被引用 7 次
- "It Brought the Model to Life": Exploring the Embodiment of Multimodal I3Ms for People who are Blind or have Low VisionSamuel Reinders, Matthew Butler, Kim MarriottCHI 2025 · 被引用 6 次
- Füpop: "Real Food" Flavor Delivery via Focused UltrasoundKatherine W. Song, Szu Ting Tung, Alexis Kim, Eric PaulosCHI 2024 · 被引用 4 次
相关 Paper
- Meta-Speaker: Acoustic Source Projection by Exploiting Air NonlinearityWeiguo Wang, Yuan He, Meng Jin, Yimiao Sun 等MobiCom 2023 · 被引用 16 次
- EarArray: Defending against DolphinAttack via Acoustic AttenuationGuoming Zhang, Xiaoyu Ji, Xinfeng Li, Gang Qu 等NDSS 2021
- Soundr: Head Position and Orientation Prediction Using a Microphone ArrayJackie (Junrui) Yang, Gaurab Banerjee, Vishesh Gupta, Monica S. Lam 等CHI 2020 · 被引用 19 次
- What If Conversational Agents Became Invisible?: Comparing Users' Mental Models According to Physical Entity of AI SpeakerSunok Lee, Minji Cho, Sangsu LeeUbiComp 2020 · 被引用 23 次
- MuDiS: An Audio-independent, Wide-angle, and Leak-free Multi-directional SpeakerYijie Li, Juntao Zhou, Dian Ding, Yi-Chao Chen 等MobiCom 2024 · 被引用 9 次
