HACK: Learning a Parametric Head and Neck Model for High-fidelity Animation
Longwen Zhang, Zijun Zhao, Xinzhou Cong, Qixuan Zhang, Shuqi Gu, Yuchong Gao, Rui Zheng, Wei Yang, Lan Xu, Jingyi Yu
Abstract
Significant advancements have been made in developing parametric models for digital humans, with various approaches concentrating on parts such as the human body, hand, or face. Nevertheless, connectors such as the neck have been overlooked in these models, with rich anatomical priors often unutilized. In this paper, we introduce HACK (Head-And-neCK), a novel parametric model for constructing the head and cervical region of digital humans. Our model seeks to disentangle the full spectrum of neck and larynx motions, facial expressions, and appearance variations, providing personalized and anatomically consistent controls, particularly for the neck regions. To build our HACK model, we acquire a comprehensive multi-modal dataset of the head and neck under various facial expressions. We employ a 3D ultrasound imaging scheme to extract the inner biomechanical structures, namely the precise 3D rotation information of the seven vertebrae of the cervical spine. We then adopt a multi-view photometric approach to capture the geometry and physically-based textures of diverse subjects, who exhibit a diverse range of static expressions as well as sequential head-and-neck movements. Using the multi-modal dataset, we train the parametric HACK model by separating the 3D head and neck depiction into various shape, pose, expression, and larynx blendshapes from the neutral expression and the rest skeletal pose. We adopt an anatomically-consistent skeletal design for the cervical region, and the expression is linked to facial action units for artist-friendly controls. We also propose to optimize the mapping from the identical shape space to the PCA spaces of personalized blendshapes to augment the pose and expression blendshapes, providing personalized properties within the framework of the generic model. Furthermore, we use larynx blendshapes to accurately control the larynx deformation and force the larynx slicing motions along the vertical direction in the UV-space for precise modeling of the larynx beneath the neck skin. HACK addresses the head and neck as a unified entity, offering more accurate and expressive controls, with a new level of realism, particularly for the neck regions. This approach has significant benefits for numerous applications, including geometric fitting and animation, and enables inter-correlation analysis between head and neck for fine-grained motion synthesis and transfer.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers6
- FlashAvatar: High-Fidelity Head Avatar with Efficient Gaussian EmbeddingJun Xiang, Xuan Gao, Yudong Guo, Juyong ZhangCVPR 2024 · 51 citations
- TGA: True-to-Geometry Avatar Dynamic ReconstructionBo Guo, Sijia Wen, Ziwei Wang, Yifan ZhaoNeurIPS 2025 · 1 citation
- WildCap: Facial Albedo Capture in the Wild via Hybrid Inverse RenderingYuxuan Han, Xin Ming, Tianxiao Li, Zhuofan Shen et al.CVPR 2026 · 1 citation
- MonoNPHM: Dynamic Head Reconstruction from Monocular VideosSimon Giebenhain, Tobias Kirschstein, Markos Georgopoulos, Martin Rünz et al.CVPR 2024
- Densemarks: Learning Canonical Embeddings for Human Heads Images via Point TracksDmitrii Pozdeev, Alexey Artemov, Ananta R. Bhattarai, Artem SevastopolskyICLR 2026
Builds on18
- Learning an animatable detailed 3D face model from in-the-wild imagesYao Feng, Haiwen Feng, Michael J. Black, Timo BolkartSIGGRAPH 2021 · 662 citations
- Mixture of volumetric primitives for efficient neural renderingStephen Lombardi, Tomas Simon, Gabriel Schwartz, Michael Zollhöfer et al.SIGGRAPH 2021 · 240 citations
- FaceFormer: Speech-Driven 3D Facial Animation with TransformersYingruo Fan, Zhaojiang Lin, Jun Saito, Wenping Wang et al.CVPR 2022 · 218 citations
- HeadNeRF: A Realtime NeRF-based Parametric Head ModelYang Hong, Bo Peng, Haiyao Xiao, Ligang Liu et al.CVPR 2022 · 189 citations
- EMOCA: Emotion Driven Monocular Face Capture and AnimationRadek Danecek, Michael J. Black, Timo BolkartCVPR 2022 · 180 citations
Related papers
- VariTex: Variational Neural Face TexturesMarcel C. Bühler, Abhimitra Meka, Gengyan Li, Thabo Beeler et al.ICCV 2021 · 43 citations
- Anatomically Detailed Simulation of Human TorsoSeunghwan Lee, Yifeng Jiang, C. Karen LiuSIGGRAPH 2023 · 1 citation
- GANHead: Towards Generative Animatable Neural Head AvatarsSijing Wu, Yichao Yan, Yunhao Li, Yuhao Cheng et al.CVPR 2023
- FaceTalk: Audio-Driven Motion Diffusion for Neural Parametric Head ModelsShivangi Aneja, Justus Thies, Angela Dai, Matthias NießnerCVPR 2024
- GHUM & GHUML: Generative 3D Human Shape and Articulated Pose ModelsHongyi Xu, Eduard Gabriel Bazavan, Andrei Zanfir, William T. Freeman et al.CVPR 2020
