PuppeteerGAN: Arbitrary Portrait Animation With Semantic-Aware Appearance Transformation
Zhuo Chen, Chaoyue Wang, Bo Yuan, Dacheng Tao
摘要
Portrait animation, which aims to animate a still portrait to life using poses extracted from target frames, is an important technique for many real-world entertainment applications. Although recent works have achieved highly realistic results on synthesizing or controlling human head images, the puppeteering of arbitrary portraits is still confronted by the following challenges: 1) identity/personality mismatch; 2) training data/domain limitations; and 3) low-efficiency in training/fine-tuning. In this paper, we devised a novel two-stage framework called PuppeteerGAN for solving these challenges. Specifically, we first learn identity-preserved semantic segmentation animation which executes pose retargeting between any portraits. As a general representation, the semantic segmentation results could be adapted to different datasets, environmental conditions or appearance domains. Furthermore, the synthesized semantic segmentation is filled with the appearance of the source portrait. To this end, an appearance transformation network is presented to produce fidelity output by jointly considering the wrapping of semantic features and conditional generation. After training, the two networks can directly perform end-to-end inference on unseen subjects without any retraining or fine-tuning. Extensive experiments on cross-identity/domain/resolution situations demonstrate the superiority of the proposed PuppetterGAN over existing portrait animation methods in both generation quality and inference speed.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper16
- PIRenderer: Controllable Portrait Image Generation via Semantic Neural RenderingYurui Ren, Ge Li, Yuanqi Chen, Thomas H. Li 等ICCV 2021 · 被引用 284 次
- SynFace: Face Recognition with Synthetic DataHaibo Qiu, Baosheng Yu, Dihong Gong, Zhifeng Li 等ICCV 2021 · 被引用 162 次
- EAMM: One-Shot Emotional Talking Face via Audio-Based Emotion-Aware Motion ModelXinya Ji, Hang Zhou, Kaisiyuan Wang, Qianyi Wu 等SIGGRAPH 2022 · 被引用 150 次
- AvatarMAV: Fast 3D Head Avatar Reconstruction Using Motion-Aware Neural VoxelsYuelang Xu, Lizhen Wang, Xiaochen Zhao, Hongwen Zhang 等SIGGRAPH 2023 · 被引用 69 次
- LatentAvatar: Learning Latent Expression Code for Expressive Neural Head AvatarYuelang Xu, Hongwen Zhang, Lizhen Wang, Xiaochen Zhao 等SIGGRAPH 2023 · 被引用 40 次
它引用的顶会 Paper4
- FSGAN: Subject Agnostic Face Swapping and ReenactmentYuval Nirkin, Yosi Keller, Tal HassnerICCV 2019 · 被引用 710 次
- Few-Shot Adversarial Learning of Realistic Neural Talking Head ModelsEgor Zakharov, Aliaksandra Shysheya, Egor Burkov, Victor S. LempitskyICCV 2019 · 被引用 687 次
- Progressive Reconstruction of Visual Structure for Image InpaintingJingyuan Li, Fengxiang He, Lefei Zhang, Bo Du 等ICCV 2019 · 被引用 151 次
- MaskGAN: Towards Diverse and Interactive Facial Image ManipulationCheng-Han Lee, Ziwei Liu, Lingyun Wu, Ping LuoCVPR 2020
相关 Paper
- Single-Shot Freestyle Dance ReenactmentOran Gafni, Oron Ashual, Lior WolfCVPR 2021
- DeX-Portrait: Disentangled and Expressive Portrait Animation via Explicit and Latent Motion RepresentationsYuxiang Shi, Zhe Li, Yanwen Wang, Hao Zhu 等CVPR 2026 · 被引用 3 次
- X-Portrait: Expressive Portrait Animation with Hierarchical Motion AttentionYou Xie, Hongyi Xu, Guoxian Song, Chao Wang 等SIGGRAPH 2024 · 被引用 40 次
- MagicPose: Realistic Human Poses and Facial Expressions Retargeting with Identity-aware DiffusionDi Chang, Yichun Shi, Quankai Gao, Hongyi Xu 等ICML 2024 · 被引用 125 次
- Durian: Dual Reference Image-Guided Portrait Animation with Attribute TransferHyunsoo Cha, Byungjun Kim, Hanbyul JooICLR 2026 · 被引用 2 次
