3D-Aware Face Editing via Warping-Guided Latent Direction Learning
Yuhao Cheng, Zhuo Chen, Xingyu Ren, Wenhan Zhu, Zhengqin Xu, Di Xu, Changpeng Yang, Yichao Yan
Abstract
3D facial editing, a longstanding task in computer vision with broad applications, is expected to fast and intuitively manipulate any face from arbitrary viewpoints following the user's will. Existing works have limitations in terms of intuitiveness, generalization, and efficiency. To overcome these challenges, we propose FaceEdit3D, which allows users to directly manipulate 3D points to edit a 3D face, achieving natural and rapid face editing. After one or several points are manipulated by users, we propose the tri-plane warping to directly deform the view-independent 3D representation. To address the problem of distortion caused by tri-plane warping, we train a warp-aware encoder to project the warped face onto a standardized latent space. In this space, we further propose directional latent editing to mitigate the identity bias caused by the encoder and realize the disentangled editing of various attributes. Extensive experiments show that our method achieves superior results with rich facial details and nice identity preservation. Our approach also supports general applications like multi-attribute continuous editing and cat/car editing. The project website is https://cyh-sj.github.io/FaceEdit3DI.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext ca14c4b8-a674-41ed-b151-8aec8fabe88bCited by top-tier papers2
- LaTo: Landmark-tokenized Diffusion Transformer for Fine-grained Human Face EditingZhenghao Zhang, Ziying Zhang, Junchao Liao, Xiangyu Meng et al.ICLR 2026 · 5 citations
- FFaceNeRF: Few-shot Face Editing in Neural Radiance FieldsKwan Yun, Chaelin Kim, Hangyeul Shin, Junyong NohCVPR 2025
Builds on38
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu et al.ICCV 2021 · 31,683 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- StyleCLIP: Text-Driven Manipulation of StyleGAN ImageryOr Patashnik, Zongze Wu, Eli Shechtman, Daniel Cohen-Or et al.ICCV 2021 · 1,437 citations
- GANSpace: Discovering Interpretable GAN ControlsErik Härkönen, Aaron Hertzmann, Jaakko Lehtinen, Sylvain ParisNeurIPS 2020 · 1,049 citations
Related papers
- AttriHuman-3D: Editable 3D Human Avatar Generation with Attribute Decomposition and IndexingFan Yang, Tianyi Chen, Xiaosheng He, Zhongang Cai et al.CVPR 2024 · 6 citations
- 3D-Aware Face SwappingYixuan Li, Chao Ma, Yichao Yan, Wenhan Zhu et al.CVPR 2023
- A Latent Transformer for Disentangled Face Editing in Images and VideosXu Yao, Alasdair Newson, Yann Gousseau, Pierre HellierICCV 2021 · 97 citations
- Reference-Based 3D-Aware Image Editing with TriplanesBahri Batuhan Bilecen, Yigit Yalin, Ning Yu, Aysegul DundarCVPR 2025
- PERSE: Personalized 3D Generative Avatars from A Single PortraitHyunsoo Cha, Inhee Lee, Hanbyul JooCVPR 2025
