Digital Ventriloquism: Giving Voice to Everyday Objects
Yasha Iravantchi, Mayank Goel, Chris Harrison
Abstract
Smart speakers with voice agents are becoming increasingly common. However, the agent's voice always emanates from the device, even when that information is contextually and spatially relevant elsewhere. Digital Ventriloquism allows smart speakers to render sound onto everyday objects, such that it appears they are speaking and are interactive. This can be achieved without any modification of objects or the environment. For this, we used a highly directional pan-tilt ultrasonic array. By modulating a 40 kHz ultrasonic signal, we can emit sound that is inaudible "in flight" and demodulates to audible frequencies when impacting a surface through acoustic parametric interaction. This makes it appear as though the sound originates from an object and not the speaker. We ran a study in which we projected speech onto five objects in three environments, and found that participants were able to correctly identify the source object 92% of the time and correctly repeat the spoken message 100% of the time, demonstrating our digital ventriloquy is both directional and intelligible.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 52d0ee70-0bcb-442e-9a2e-19b674ee0500Cited by top-tier papers6
- Evoking empathy with visually impaired people through an augmented reality embodiment experienceRenan Guarese, Emma Pretty, Haytham M. Fayek, Fabio Zambetta et al.IEEE VR 2023 · 23 citations
- Auptimize: Optimal Placement of Spatial Audio Cues for Extended RealityHyunsung Cho, Alexander Wang, Divya Kartik, Emily Liying Xie et al.UIST 2024 · 19 citations
- Bubbleu: Exploring Augmented Reality Game Design with Uncertain AI-based InteractionMinji Kim, Kyungjin Lee, Rajesh Balan, Youngki LeeCHI 2023 · 7 citations
- "It Brought the Model to Life": Exploring the Embodiment of Multimodal I3Ms for People who are Blind or have Low VisionSamuel Reinders, Matthew Butler, Kim MarriottCHI 2025 · 6 citations
- Füpop: "Real Food" Flavor Delivery via Focused UltrasoundKatherine W. Song, Szu Ting Tung, Alexis Kim, Eric PaulosCHI 2024 · 4 citations
Related papers
- Meta-Speaker: Acoustic Source Projection by Exploiting Air NonlinearityWeiguo Wang, Yuan He, Meng Jin, Yimiao Sun et al.MobiCom 2023 · 16 citations
- EarArray: Defending against DolphinAttack via Acoustic AttenuationGuoming Zhang, Xiaoyu Ji, Xinfeng Li, Gang Qu et al.NDSS 2021
- Soundr: Head Position and Orientation Prediction Using a Microphone ArrayJackie (Junrui) Yang, Gaurab Banerjee, Vishesh Gupta, Monica S. Lam et al.CHI 2020 · 19 citations
- What If Conversational Agents Became Invisible?: Comparing Users' Mental Models According to Physical Entity of AI SpeakerSunok Lee, Minji Cho, Sangsu LeeUbiComp 2020 · 23 citations
- MuDiS: An Audio-independent, Wide-angle, and Leak-free Multi-directional SpeakerYijie Li, Juntao Zhou, Dian Ding, Yi-Chao Chen et al.MobiCom 2024 · 9 citations
