Make a Face: Towards Arbitrary High Fidelity Face Manipulation
Shengju Qian, Kwan-Yee Lin, Wayne Wu, Yangxiaokang Liu, Quan Wang, Fumin Shen, Chen Qian, Ran He
Abstract
Recent studies have shown remarkable success in face manipulation task with the advance of GANs and VAEs paradigms, but the outputs are sometimes limited to low-resolution and lack of diversity. In this work, we propose Additive Focal Variational Auto-encoder (AF-VAE), a novel approach that can arbitrarily manipulate high-resolution face images using a simple yet effective model and only weak supervision of reconstruction and KL divergence losses. First, a novel additive Gaussian Mixture assumption is introduced with an unsupervised clustering mechanism in the structural latent space, which endows better disentanglement and boosts multi-modal representation with external memory. Second, to improve the perceptual quality of synthesized results, two simple strategies in architecture design are further tailored and discussed on the behavior of Human Visual System (HVS) for the first time, allowing for fine control over the model complexity and sample quality. Human opinion studies and new state-of-the-art Inception Score (IS) / Frechet Inception Distance (FID) demonstrate the superiority of our approach over existing algorithms, advancing both the fidelity and extremity of face manipulation task.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 094396cc-0654-4f9c-a09b-747b57df88c3Cited by top-tier papers12
- MagicAnimate: Temporally Consistent Human Image Animation using Diffusion ModelZhongcong Xu, Jianfeng Zhang, Jun Hao Liew, Hanshu Yan et al.CVPR 2024 · 106 citations
- Controllable 3D Face Synthesis with Conditional Generative Occupancy FieldsKeqiang Sun, Shangzhe Wu, Zhaoyang Huang, Ning Zhang et al.NeurIPS 2022 · 62 citations
- Re-Aging GAN: Toward Personalized Face Age TransformationFarkhod Makhmudkhujaev, Sungeun Hong, In Kyu ParkICCV 2021 · 35 citations
- Harnessing the Conditioning Sensorium for Improved Image TranslationCooper Nederhood, Nicholas I. Kolkin, Deqing Fu, Jason SalavonICCV 2021 · 6 citations
- Canonswap: High-Fidelity and Consistent Video Face Swapping Via Canonical Space ModulationXiangyang Luo, Ye Zhu, Yunfei Liu, Lijian Lin et al.ICCV 2025 · 4 citations
Related papers
- Interpreting the Latent Space of GANs for Semantic Face EditingYujun Shen, Jinjin Gu, Xiaoou Tang, Bolei ZhouCVPR 2020
- SDGAN: Disentangling Semantic Manipulation for Facial Attribute EditingWenmin Huang, Weiqi Luo, Jiwu Huang, Xiaochun CaoAAAI 2024 · 20 citations
- D2Animator: Dual Distillation of StyleGAN For High-Resolution Face AnimationZhuo Chen, Chaoyue Wang, Haimei Zhao, Bo Yuan et al.ACM MM 2022 · 3 citations
- Vec2Face: Scaling Face Dataset Generation with Loosely Constrained VectorsHaiyu Wu, Jaskirat Singh, Sicong Tian, Liang Zheng et al.ICLR 2025
- FaceController: Controllable Attribute Editing for Face in the WildZhiliang Xu, Xiyu Yu, Zhibin Hong, Zhen Zhu et al.AAAI 2021 · 49 citations
