Model-based Head Orientation Estimation for Smart Devices
Qiang Yang, Yuanqing Zheng
摘要
Voice interaction is friendly and convenient for users. Smart devices such as Amazon Echo allow users to interact with them by voice commands and become increasingly popular in our daily life. In recent years, research works focus on using the microphone array built in smart devices to localize the user's position, which adds additional context information to voice commands. In contrast, few works explore the user's head orientation, which also contains useful context information. For example, when a user says, "turn on the light", the head orientation could infer which light the user is referring to. Existing model-based works require a large number of microphone arrays to form an array network, while machine learning-based approaches need laborious data collection and training workload. High deployment/usage cost of these methods is unfriendly to users. In this paper, we propose HOE, a model-based system that enables Head Orientation Estimation for smart devices with only two microphone arrays, which requires a lower training overhead than previous approaches. HOE first estimates the user's head orientation candidates by measuring the voice energy radiation pattern. Then, the voice frequency radiation pattern is leveraged to obtain the final result. Real-world experiments are conducted, and the results show that HOE can achieve a median estimation error of 23 degrees. To the best of our knowledge, HOE is the first model-based attempt to estimate the head orientation by only two microphone arrays without the arduous data training overhead.
CCS Concepts: • Human-centered computing → Ubiquitous and mobile computing design and evaluation methods.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- VoShield: Voice Liveness Detection with Sound Field DynamicsQiang Yang, Kaiyan Cui, Yuanqing ZhengINFOCOM 2023 · 被引用 13 次
- HearFire: Indoor Fire Detection via Inaudible Acoustic SensingZheng Wang, Yanwen Wang, Mi Tian, Jiaxing ShenUbiComp 2023 · 被引用 8 次
- Facial Landmark Detection Based on High Precision Spatial Sampling via Millimeter-wave RadarYi Li, Chuyu Wang, Lei Xie, Qiancheng Jin 等UbiComp 2025 · 被引用 7 次
它引用的顶会 Paper3
- Voice localization using nearby wall reflectionsSheng Shen, Daguan Chen, Yu-Lin Wei, Zhijian Yang 等MobiCom 2020 · 被引用 79 次
- Soundr: Head Position and Orientation Prediction Using a Microphone ArrayJackie (Junrui) Yang, Gaurab Banerjee, Vishesh Gupta, Monica S. Lam 等CHI 2020 · 被引用 19 次
- AcouRadar: Towards Single Source based Acoustic LocalizationLinsong Cheng, Zhao Wang, Yunting Zhang, Weiyi Wang 等INFOCOM 2020 · 被引用 13 次
相关 Paper
- Direction-of-Voice (DoV) Estimation for Intuitive Speech Interaction with Smart Devices EcosystemsKaran Ahuja, Andy Kong, Mayank Goel, Chris HarrisonUIST 2020 · 被引用 29 次
- MAVL: Multiresolution Analysis of Voice LocalizationMei Wang, Wei Sun, Lili QiuNSDI 2021 · 被引用 46 次
- EarArray: Defending against DolphinAttack via Acoustic AttenuationGuoming Zhang, Xiaoyu Ji, Xinfeng Li, Gang Qu 等NDSS 2021
- FaceOri: Tracking Head Position and Orientation Using Ultrasonic Ranging on EarphonesYuntao Wang, Jiexin Ding, Ishan Chatterjee, Farshid Salemi Parizi 等CHI 2022 · 被引用 29 次
- Enhancing Mobile Voice Assistants with WorldGazeSven Mayer, Gierad Laput, Chris HarrisonCHI 2020 · 被引用 65 次
