DipGuava: Disentangling Personalized Gaussian Features for 3D Head Avatars from Monocular Video
Jeonghaeng Lee, Seokkeun Choi, Zhixuan Li, Weisi Lin, Sanghoon Lee
摘要
While recent 3D head avatar creation methods attempt to animate facial dynamics, they often fail to capture personalized details, limiting realism and expressiveness. To fill this gap, we present DipGuava (Disentangled and Personalized Gaussian UV Avatar), a novel 3D Gaussian head avatar creation method that successfully generates avatars with personalized attributes from monocular video. DipGuava is the first method to explicitly disentangle facial appearance into two complementary components, trained in a structured two-stage pipeline that significantly reduces learning ambiguity and enhances reconstruction fidelity. In the first stage, we learn a stable geometry-driven base appearance that captures global facial structure and coarse expression-dependent variations. In the second stage, the personalized residual details not captured in the first stage are predicted, including high-frequency components and nonlinearly varying features such as wrinkles and subtle skin deformations. These components are fused via dynamic appearance fusion that integrates residual details after deformation, ensuring spatial and semantic alignment. This disentangled design enables DipGuava to generate photorealistic, identity-preserving avatars, consistently outperforming prior methods in both visual quality and quantitative performance, as demonstrated in extensive experiments.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper20
- 3D Gaussian Splatting for Real-Time Radiance Field RenderingBernhard Kerbl, Georgios Kopanas, Thomas Leimkühler, George DrettakisSIGGRAPH 2023 · 被引用 5,687 次
- HeadNeRF: A Realtime NeRF-based Parametric Head ModelYang Hong, Bo Peng, Haiyao Xiao, Ligang Liu 等CVPR 2022 · 被引用 189 次
- GaussianAvatars: Photorealistic Head Avatars with Rigged 3D GaussiansShenhan Qian, Tobias Kirschstein, Liam Schoneveld, Davide Davoli 等CVPR 2024 · 被引用 175 次
- Neural Head Avatars from Monocular RGB VideosPhilip-William Grassal, Malte Prinzler, Titus Leistner, Carsten Rother 等CVPR 2022 · 被引用 173 次
- I M Avatar: Implicit Morphable Head Avatars from VideosYufeng Zheng, Victoria Fernández Abrevaya, Marcel C. Bühler, Xu Chen 等CVPR 2022 · 被引用 169 次
相关 Paper
- GaussianAvatar: Towards Realistic Human Avatar Modeling from a Single Video via Animatable 3D GaussiansLiangxiao Hu, Hongwen Zhang, Yuxiang Zhang, Boyao Zhou 等CVPR 2024
- 3D Gaussian Blendshapes for Head Avatar AnimationShengjie Ma, Yanlin Weng, Tianjia Shao, Kun ZhouSIGGRAPH 2024 · 被引用 57 次
- FlashAvatar: High-Fidelity Head Avatar with Efficient Gaussian EmbeddingJun Xiang, Xuan Gao, Yudong Guo, Juyong ZhangCVPR 2024 · 被引用 51 次
- TeGA: Texture Space Gaussian Avatars for High-Resolution Dynamic Head ModelingGengyan Li, Paulo F. U. Gotardo, Timo Bolkart, Stephan J. Garbin 等SIGGRAPH 2025 · 被引用 2 次
- MonoGaussianAvatar: Monocular Gaussian Point-based Head AvatarYufan Chen, Lizhen Wang, Qijing Li, Hongjiang Xiao 等SIGGRAPH 2024 · 被引用 85 次
