Make a Face: Towards Arbitrary High Fidelity Face Manipulation
Shengju Qian, Kwan-Yee Lin, Wayne Wu, Yangxiaokang Liu, Quan Wang, Fumin Shen, Chen Qian, Ran He
摘要
Recent studies have shown remarkable success in face manipulation task with the advance of GANs and VAEs paradigms, but the outputs are sometimes limited to low-resolution and lack of diversity. In this work, we propose Additive Focal Variational Auto-encoder (AF-VAE), a novel approach that can arbitrarily manipulate high-resolution face images using a simple yet effective model and only weak supervision of reconstruction and KL divergence losses. First, a novel additive Gaussian Mixture assumption is introduced with an unsupervised clustering mechanism in the structural latent space, which endows better disentanglement and boosts multi-modal representation with external memory. Second, to improve the perceptual quality of synthesized results, two simple strategies in architecture design are further tailored and discussed on the behavior of Human Visual System (HVS) for the first time, allowing for fine control over the model complexity and sample quality. Human opinion studies and new state-of-the-art Inception Score (IS) / Frechet Inception Distance (FID) demonstrate the superiority of our approach over existing algorithms, advancing both the fidelity and extremity of face manipulation task.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper12
- MagicAnimate: Temporally Consistent Human Image Animation using Diffusion ModelZhongcong Xu, Jianfeng Zhang, Jun Hao Liew, Hanshu Yan 等CVPR 2024 · 被引用 106 次
- Controllable 3D Face Synthesis with Conditional Generative Occupancy FieldsKeqiang Sun, Shangzhe Wu, Zhaoyang Huang, Ning Zhang 等NeurIPS 2022 · 被引用 62 次
- Re-Aging GAN: Toward Personalized Face Age TransformationFarkhod Makhmudkhujaev, Sungeun Hong, In Kyu ParkICCV 2021 · 被引用 35 次
- Harnessing the Conditioning Sensorium for Improved Image TranslationCooper Nederhood, Nicholas I. Kolkin, Deqing Fu, Jason SalavonICCV 2021 · 被引用 6 次
- Canonswap: High-Fidelity and Consistent Video Face Swapping Via Canonical Space ModulationXiangyang Luo, Ye Zhu, Yunfei Liu, Lijian Lin 等ICCV 2025 · 被引用 4 次
相关 Paper
- Interpreting the Latent Space of GANs for Semantic Face EditingYujun Shen, Jinjin Gu, Xiaoou Tang, Bolei ZhouCVPR 2020
- SDGAN: Disentangling Semantic Manipulation for Facial Attribute EditingWenmin Huang, Weiqi Luo, Jiwu Huang, Xiaochun CaoAAAI 2024 · 被引用 20 次
- D2Animator: Dual Distillation of StyleGAN For High-Resolution Face AnimationZhuo Chen, Chaoyue Wang, Haimei Zhao, Bo Yuan 等ACM MM 2022 · 被引用 3 次
- Vec2Face: Scaling Face Dataset Generation with Loosely Constrained VectorsHaiyu Wu, Jaskirat Singh, Sicong Tian, Liang Zheng 等ICLR 2025
- FaceController: Controllable Attribute Editing for Face in the WildZhiliang Xu, Xiyu Yu, Zhibin Hong, Zhen Zhu 等AAAI 2021 · 被引用 49 次
