Capturing the Unseen: Vision-Free Facial Motion Capture Using Inertial Measurement Units
Youjia Wang, Yiwen Wu, Hengan Zhou, Hongyang Lin, Xingyue Peng, Jingyan Zhang, Yingsheng Zhu, Yingwenqi Jiang, Yatu Zhang, Lan Xu, Jingya Wang, Jingyi Yu
摘要
We present Capturing the Unseen (CAPUS), a novel facial motion capture (MoCap) technique that operates without visual signals. CAPUS leverages miniaturized Inertial Measurement Units (IMUs) as a new sensing modality for facial motion capture. While IMUs have become essential in fullbody MoCap for their portability and independence from environmental conditions, their application in facial MoCap remains underexplored. We address this by customizing micro-IMUs, small enough to be placed on the face, and strategically positioning them in alignment with key facial muscles to capture expression dynamics. CAPUS introduces the first facial IMU dataset, encompassing both IMU and visual signals from participants engaged in diverse activities such as multilingual speech, facial expressions, and emotionally intoned auditions. We train a Transformer Diffusion-based neural network to infer Blendshape parameters directly from IMU data. Our experimental results demonstrate that CAPUS reliably captures facial motion in conditions where visual-based methods struggle, including facial occlusions, rapid movements, and low-light environments. Additionally, by eliminating the need for visual inputs, CAPUS offers enhanced privacy protection, making it a robust solution for vision-free facial MoCap.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper6
- Learning an animatable detailed 3D face model from in-the-wild imagesYao Feng, Haiwen Feng, Michael J. Black, Timo BolkartSIGGRAPH 2021 · 被引用 662 次
- TransPose: real-time 3D human translation and pose estimation with six inertial sensorsXinyu Yi, Yuxiao Zhou, Feng XuSIGGRAPH 2021 · 被引用 200 次
- DreamFace: Progressive Generation of Animatable 3D Faces under Text GuidanceLongwen Zhang, Qiwei Qiu, Hongyang Lin, Qixuan Zhang 等SIGGRAPH 2023 · 被引用 68 次
- A Morphable Face Albedo ModelWilliam A. P. Smith, Alassane Seck, Hannah M. Dee, Bernard Tiddeman 等CVPR 2020
- Learning Formation of Physically-Based Face AttributesRuilong Li, Karl Bladin, Yajie Zhao, Chinmay Chinara 等CVPR 2020
相关 Paper
- ExpressEar: Sensing Fine-Grained Facial Expressions with EarablesDhruv Verma, Sejal Bhalla, Dhruv Sahnan, Jainendra Shukla 等UbiComp 2021 · 被引用 59 次
- Sensor-Augmented Egocentric-Video Captioning with Dynamic Modal AttentionKatsuyuki Nakamura, Hiroki Ohashi, Mitsuhiro OkadaACM MM 2021 · 被引用 9 次
- IMU-HOI: A Symbiotic Framework for Coherent Human-Object Interaction and Motion Capture via Contact-Conscious Inertial FusionLizhou Lin, Songpengcheng Xia, Zengyuan Lai, Lan Sun 等CVPR 2026 · 被引用 1 次
- Synthetic Smartwatch IMU Data Generation from In-the-wild ASL VideosPanneer Selvam Santhalingam, Parth Pathak, Huzefa Rangwala, Jana KoseckaUbiComp 2023 · 被引用 28 次
- NaME: A Natural Micro-expression Dataset for Micro-expression Recognition in the WildJiateng Liu, Hengcan Shi, Haiwen Liang, Xiaolin Xu 等ACM MM 2025 · 被引用 4 次
