Generalizable One-shot 3D Neural Head Avatar
Xueting Li, Shalini De Mello, Sifei Liu, Koki Nagano, Umar Iqbal, Jan Kautz
摘要
We present a method that reconstructs and animates a 3D head avatar from a single-view portrait image. Existing methods either involve time-consuming optimization for a specific person with multiple images, or they struggle to synthesize intricate appearance details beyond the facial region. To address these limitations, we propose a framework that not only generalizes to unseen identities based on a single-view image without requiring person-specific optimization, but also captures characteristic details within and beyond the face area ( e.g . hairstyle, accessories, etc .). At the core of our method are three branches that produce three tri-planes representing the coarse 3D geometry, detailed appearance of a source image, as well as the expression of a target image. By applying volumetric rendering to the combination of the three tri-planes followed by a super-resolution module, our method yields a high fidelity image of the desired identity, expression and pose. Once trained, our model enables efficient 3D head avatar reconstruction and animation via a single forward pass through a network. Experiments show that the proposed approach generalizes well to unseen validation datasets, surpassing SOTA baseline methods by a large margin on head avatar reconstruction and animation.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper17
- GaussianTalker: Speaker-specific Talking Head Synthesis via 3D Gaussian SplattingHongyun Yu, Zhan Qu, Qihang Yu, Jianchuan Chen 等ACM MM 2024 · 被引用 28 次
- MimicTalk: Mimicking a personalized and expressive 3D talking face in minutesZhenhui Ye, Tianyun Zhong, Yi Ren, Ziyue Jiang 等NeurIPS 2024 · 被引用 28 次
- LAM: Large Avatar Model for One-shot Animatable Gaussian HeadYisheng He, Xiaodong Gu, Xiaodan Ye, Chao Xu 等SIGGRAPH 2025 · 被引用 14 次
- Dual Encoder GAN Inversion for High-Fidelity 3D Head Reconstruction from Single ImagesBahri Batuhan Bilecen, Ahmet Berke Gökmen, Aysegul DundarNeurIPS 2024 · 被引用 11 次
- FlexAvatar: Learning Complete 3D Head Avatars with Partial SupervisionTobias Kirschstein, Simon Giebenhain, Matthias NießnerCVPR 2026 · 被引用 10 次
它引用的顶会 Paper34
- SegFormer: Simple and Efficient Design for Semantic Segmentation with TransformersEnze Xie, Wenhai Wang, Zhiding Yu, Anima Anandkumar 等NeurIPS 2021 · 被引用 9,661 次
- Instant neural graphics primitives with a multiresolution hash encodingThomas Müller, Alex Evans, Christoph Schied, Alexander KellerSIGGRAPH 2022 · 被引用 4,089 次
- Efficient Geometry-aware 3D Generative Adversarial NetworksEric R. Chan, Connor Z. Lin, Matthew A. Chan, Koki Nagano 等CVPR 2022 · 被引用 984 次
- Learning an animatable detailed 3D face model from in-the-wild imagesYao Feng, Haiwen Feng, Michael J. Black, Timo BolkartSIGGRAPH 2021 · 被引用 662 次
- Non-Rigid Neural Radiance Fields: Reconstruction and Novel View Synthesis of a Dynamic Scene From Monocular VideoEdgar Tretschk, Ayush Tewari, Vladislav Golyanik, Michael Zollhöfer 等ICCV 2021 · 被引用 617 次
相关 Paper
- Real3D-Portrait: One-shot Realistic 3D Talking Portrait SynthesisZhenhui Ye, Tianyun Zhong, Yi Ren, Jiaqi Yang 等ICLR 2024 · 被引用 105 次
- NOFA: NeRF-based One-shot Facial Avatar ReconstructionWangbo Yu, Yanbo Fan, Yong Zhang, Xuan Wang 等SIGGRAPH 2023 · 被引用 38 次
- SOAP: Style-Omniscient Animatable PortraitsTingting Liao, Yujian Zheng, Yuliang Xiu, Adilbek Karmanov 等SIGGRAPH 2025 · 被引用 2 次
- VASA-3D: Lifelike Audio-Driven Gaussian Head Avatars from a Single ImageSicheng Xu, Guojun Chen, Jiaolong Yang, Yizhong Zhang 等NeurIPS 2025 · 被引用 5 次
- OMG-Avatar: One-shot Multi-LOD Gaussian Head AvatarJianqiang Ren, Lin Liu, Steven HoiCVPR 2026 · 被引用 1 次
