High-Quality Full-Head 3D Avatar Generation from Any Single Portrait Image
Yujie Gao, Chencheng Wang, Xianbing Sun, Jiahui Zhan, Wentao Wang, Yiyi Zhang, Haohua Zhao, Liqing Zhang, Jianfu Zhang
摘要
In this work, we introduce a novel high-fidelity full-head 3D avatar generation method from a single image, regardless of perspective, style, expression, or accessories. Prior works often fail to preserve consistent head geometry and facial details, primarily due to their limited capacity in modeling fine-grained facial textures and maintaining identity information. To address these challenges, we construct a new high-quality dataset containing 227 sequences of digital human portraits captured from 96 different perspectives, totaling 21,792 frames, featuring high-quality facial texture details. To further improve performance, we propose a novel multi-view diffusion model named ID-TS diffusion model, which integrates identity and expression information into the two-stage multi-view diffusion process. The low-resolution stage ensures structural consistency of heads across multiple views, while the high-resolution stage preserves facial detail fidelity and coherence. Finally, we propose an enhanced feed-forward Gaussian avatar reconstruction method that optimizes the network on multi-view images of each single subject, significantly improving 3D facial texture details. Extensive experiments demonstrate that our method achieves robust performance across challenging scenarios, while showcasing broad applicability across numerous downstream tasks.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper20
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- 3D Gaussian Splatting for Real-Time Radiance Field RenderingBernhard Kerbl, Georgios Kopanas, Thomas Leimkühler, George DrettakisSIGGRAPH 2023 · 被引用 5,687 次
- Efficient Geometry-aware 3D Generative Adversarial NetworksEric R. Chan, Connor Z. Lin, Matthew A. Chan, Koki Nagano 等CVPR 2022 · 被引用 984 次
相关 Paper
- SAMT: Generating Structured Avatar Meshes and Textures from a Single ImageMuyu Wang, Jianzhe Gao, Xingping Dong, Yujia Wang 等ICML 2026
- GeoDiff4D: Geometry-Aware Diffusion for 4D Head Avatar ReconstructionChao Xu, Xiaochen Zhao, Xiang Deng, Jingxiang Sun 等CVPR 2026
- ID-to-3D: Expressive ID-guided 3D Heads via Score Distillation SamplingFrancesca Babiloni, Alexandros Lattas, Jiankang Deng, Stefanos ZafeiriouNeurIPS 2024 · 被引用 5 次
- ConsistentAvatar: Learning to Diffuse Fully Consistent Talking Head Avatar with Temporal GuidanceHaijie Yang, Zhenyu Zhang, Hao Tang, Jianjun Qian 等ACM MM 2024 · 被引用 3 次
- MoGA: 3D Generative Avatar Prior for Monocular Gaussian Avatar ReconstructionZijian Dong, Longteng Duan, Jie Song, Michael J. Black 等ICCV 2025 · 被引用 4 次
