Theophany: Multimodal Speech Augmentation in Instantaneous Privacy Channels
Abhishek Kumar, Tristan Braud, Lik Hang Lee, Pan Hui
Abstract
Many factors affect speech intelligibility in face-to-face conversations. These factors lead conversation participants to speak louder and more distinctively, exposing the content to potential eavesdroppers. To address these issues, we introduce Theophany, a privacy-preserving framework for augmenting speech. Theophany establishes ad-hoc social networks between conversation participants to exchange contextual information, improving speech intelligibility in real-time. At the core of Theophany, we develop the first privacy perception model that assesses the privacy risk of a face-to-face conversation based on its topic, location, and participants. This framework allows to develop any privacy-preserving application for face-to-face conversation. We implement the framework within a prototype system that augments the speaker's speech with real-life subtitles to overcome the loss of contextual cues brought by mask-wearing and social distancing during the COVID-19 pandemic. We evaluate Theophany through a user survey and a user study on 53 and 17 participants, respectively. Theophany's privacy predictions match the participants' privacy preferences with an accuracy of 71.26%. Users considered Theophany to be useful to protect their privacy (3.88/5), easy to use (4.71/5), and enjoyable to use (4.24/5). We also raise the question of demographic and individual differences in the design of privacy-preserving solutions.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Builds on2
- How Subtle Can It Get?: A Trimodal Study of Ring-sized Interfaces for One-Handed Drone ControlYui-Pan Yau, Lik Hang Lee, Zheng Li, Tristan Braud et al.UbiComp 2020 · 29 citations
- Aquilis: Using Contextual Integrity for Privacy Protection on Mobile DevicesAbhishek Kumar, Tristan Braud, Young D. Kwon, Pan HuiUbiComp 2021 · 13 citations
Related papers
- Whispering Under the Eaves: Protecting User Privacy Against Commercial and LLM-powered Automatic Speech Recognition SystemsWeifei Jin, Yuxin Cao, Junjie Su, Derui Wang et al.USENIX Security 2025
- MEPS: Privacy-preserving Edge-cloud Video Foundation Model Inference with Privacy ProtectabilitySiping Shi, Rui Lu, Dan Wang, Bihai ZhangUSENIX Security 2026
- SafeSpeaker: Voice Obfuscation for Resource-Constrained IoT DevicesCameron Haire, Yasha Iravantchi, Kang G. Shin, Alanson P. SampleUbiComp 2026 · 1 citation
- AeroSense: Sensing Aerosol Emissions from Indoor Human ActivitiesBhawana Chhaglani, Camellia Zakaria, Richard Peltier, Jeremy Gummeson et al.UbiComp 2024 · 6 citations
- COMPA: Using Conversation Context to Achieve Common Ground in AACStephanie Valencia, Jessica Huynh, Emma Y. Jiang, Yufei Wu et al.CHI 2024 · 20 citations
