STAvatar: Soft Binding and Temporal Density Control for Monocular 3D Head Avatars Reconstruction
Jiankuo Zhao, Xiangyu Zhu, Zidu Wang, Zhen Lei
Abstract
Reconstructing high-fidelity and animatable 3D head avatars from monocular videos remains a challenging yet essential task. Existing methods based on 3D Gaussian Splatting typically bind Gaussians to mesh triangles and model deformations solely via Linear Blend Skinning, which results in rigid motion and limited expressiveness. Moreover, they lack specialized strategies to handle frequently occluded regions (e.g., mouth interiors, eyelids). To address these limitations, we propose STAvatar, which consists of two key components: (1) a UV-Adaptive Soft Binding framework that leverages both image-based and geometric priors to learn per-Gaussian feature offsets within the UV space. This UV representation supports dynamic resampling, ensuring full compatibility with Adaptive Density Control (ADC) and enhanced adaptability to shape and textural variations. (2) a Temporal ADC strategy, which first clusters structurally similar frames to facilitate more targeted computation of the densification criterion. It further introduces a novel fused perceptual error as clone criterion to jointly capture geometric and textural discrepancies, encouraging densification in regions requiring finer details. Extensive experiments on four benchmark datasets demonstrate that STAvatar achieves state-of-the-art reconstruction performance, especially in capturing fine-grained details and reconstructing frequently occluded regions. Our project is available at https://jiankuozhao.github.io/STAvatar/.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 8d99e3c8-c665-4d39-b6c7-79561c087448Cited by top-tier papers1
Ask how each one uses itBuilds on31
- 3D Gaussian Splatting for Real-Time Radiance Field RenderingBernhard Kerbl, Georgios Kopanas, Thomas Leimkühler, George DrettakisSIGGRAPH 2023 · 5,687 citations
- Learning an animatable detailed 3D face model from in-the-wild imagesYao Feng, Haiwen Feng, Michael J. Black, Timo BolkartSIGGRAPH 2021 · 662 citations
- 3D Gaussian Splatting as Markov Chain Monte CarloShakiba Kheradmand, Daniel Rebain, Gopal Sharma, Weiwei Sun et al.NeurIPS 2024 · 285 citations
- GaussianAvatars: Photorealistic Head Avatars with Rigged 3D GaussiansShenhan Qian, Tobias Kirschstein, Liam Schoneveld, Davide Davoli et al.CVPR 2024 · 175 citations
- Neural Head Avatars from Monocular RGB VideosPhilip-William Grassal, Malte Prinzler, Titus Leistner, Carsten Rother et al.CVPR 2022 · 173 citations
Related papers
- TeGA: Texture Space Gaussian Avatars for High-Resolution Dynamic Head ModelingGengyan Li, Paulo F. U. Gotardo, Timo Bolkart, Stephan J. Garbin et al.SIGGRAPH 2025 · 2 citations
- MonoGaussianAvatar: Monocular Gaussian Point-based Head AvatarYufan Chen, Lizhen Wang, Qijing Li, Hongjiang Xiao et al.SIGGRAPH 2024 · 85 citations
- FATE: Full-head Gaussian Avatar with Textural Editing from Monocular VideoJiawei Zhang, Zijian Wu, Zhiyang Liang, Yicheng Gong et al.CVPR 2025
- Relightable and Dynamic Gaussian Avatar Reconstruction from Monocular VideoSeonghwa Choi, Moonkyeong Choi, Mingyu Jang, Jaekyung Kim et al.ACM MM 2025 · 1 citation
- 3D Gaussian Blendshapes for Head Avatar AnimationShengjie Ma, Yanlin Weng, Tianjia Shao, Kun ZhouSIGGRAPH 2024 · 57 citations
