StyleAvatar: Real-time Photo-realistic Portrait Avatar from a Single Video
Lizhen Wang, Xiaochen Zhao, Jingxiang Sun, Yuxiang Zhang, Hongwen Zhang, Tao Yu, Yebin Liu
摘要
Face reenactment methods attempt to restore and re-animate portrait videos as realistically as possible. Existing methods face a dilemma in quality versus controllability: 2D GAN-based methods achieve higher image quality but suffer in fine-grained control of facial attributes compared with 3D counterparts. In this work, we propose StyleAvatar, a real-time photo-realistic portrait avatar reconstruction method using StyleGAN-based networks, which can generate high-fidelity portrait avatars with faithful expression control. We expand the capabilities of StyleGAN by introducing a compositional representation and a sliding window augmentation method, which enable faster convergence and improve translation generalization. Specifically, we divide the portrait scenes into three parts for adaptive adjustments: facial region, non-facial foreground region, and the background. Besides, our network leverages the best of UNet, StyleGAN and time coding for video learning, which enables high-quality video generation. Furthermore, a sliding window augmentation method together with a pre-training strategy are proposed to improve translation generalization and training performance, respectively. The proposed network can converge within two hours while ensuring high image quality and a forward rendering time of only 20 milliseconds. Furthermore, we propose a real-time live system, which further pushes research into applications. Results and experiments demonstrate the superiority of our method in terms of image quality, full portrait video generation, and real-time re-animation compared to existing facial reenactment methods. Training and inference code for this paper are at https://github.com/LizhenWangT/StyleAvatar.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper24
- AvatarReX: Real-time Expressive Full-body AvatarsZerong Zheng, Xiaochen Zhao, Hongwen Zhang, Boning Liu 等SIGGRAPH 2023 · 被引用 80 次
- AvatarVerse: High-Quality & Stable 3D Avatar Creation from Text and PoseHuichao Zhang, Bowen Chen, Hao Yang, Liao Qu 等AAAI 2024 · 被引用 73 次
- AvatarMAV: Fast 3D Head Avatar Reconstruction Using Motion-Aware Neural VoxelsYuelang Xu, Lizhen Wang, Xiaochen Zhao, Hongwen Zhang 等SIGGRAPH 2023 · 被引用 69 次
- LatentAvatar: Learning Latent Expression Code for Expressive Neural Head AvatarYuelang Xu, Hongwen Zhang, Lizhen Wang, Xiaochen Zhao 等SIGGRAPH 2023 · 被引用 40 次
- LayGA: Layered Gaussian Avatars for Animatable Clothing TransferSiyou Lin, Zhe Li, Zhaoqi Su, Zerong Zheng 等SIGGRAPH 2024 · 被引用 27 次
它引用的顶会 Paper27
- Instant neural graphics primitives with a multiresolution hash encodingThomas Müller, Alex Evans, Christoph Schied, Alexander KellerSIGGRAPH 2022 · 被引用 4,089 次
- Alias-Free Generative Adversarial NetworksTero Karras, Miika Aittala, Samuli Laine, Erik Härkönen 等NeurIPS 2021 · 被引用 2,126 次
- GANSpace: Discovering Interpretable GAN ControlsErik Härkönen, Aaron Hertzmann, Jaakko Lehtinen, Sylvain ParisNeurIPS 2020 · 被引用 1,049 次
- Efficient Geometry-aware 3D Generative Adversarial NetworksEric R. Chan, Connor Z. Lin, Matthew A. Chan, Koki Nagano 等CVPR 2022 · 被引用 984 次
- ReStyle: A Residual-Based StyleGAN Encoder via Iterative RefinementYuval Alaluf, Or Patashnik, Daniel Cohen-OrICCV 2021 · 被引用 377 次
相关 Paper
- Unsupervised Facial Performance Editing via Vector-Quantized StyleGAN RepresentationsBerkay Kicanaoglu, Pablo Garrido, Gaurav BharajICCV 2023 · 被引用 2 次
- Normalized Avatar Synthesis Using StyleGAN and Perceptual RefinementHuiwen Luo, Koki Nagano, Han-Wei Kung, Qingguo Xu 等CVPR 2021
- StyleRig: Rigging StyleGAN for 3D Control Over Portrait ImagesAyush Tewari, Mohamed A. Elgharib, Gaurav Bharaj, Florian Bernard 等CVPR 2020
- AgileGAN: stylizing portraits by inversion-consistent transfer learningGuoxian Song, Linjie Luo, Jing Liu, Wan-Chun Ma 等SIGGRAPH 2021 · 被引用 80 次
- DeepFaceVideoEditing: sketch-based deep editing of face videosFeng-Lin Liu, Shu-Yu Chen, Yu-Kun Lai, Chunpeng Li 等SIGGRAPH 2022 · 被引用 26 次
