DeepFaceVideoEditing: sketch-based deep editing of face videos
Feng-Lin Liu, Shu-Yu Chen, Yu-Kun Lai, Chunpeng Li, Yue-Ren Jiang, Hongbo Fu, Lin Gao
Abstract
Sketches, which are simple and concise, have been used in recent deep image synthesis methods to allow intuitive generation and editing of facial images. However, it is nontrivial to extend such methods to video editing due to various challenges, ranging from appropriate manipulation propagation and fusion of multiple editing operations to ensure temporal coherence and visual quality. To address these issues, we propose a novel sketch-based facial video editing framework, in which we represent editing manipulations in latent space and propose specific propagation and fusion modules to generate high-quality video editing results based on StyleGAN3. Specifically, we first design an optimization approach to represent sketch editing manipulations by editing vectors, which are propagated to the whole video sequence using a proper strategy to cope with different editing needs. Specifically, input editing operations are classified into two categories: temporally consistent editing and temporally variant editing. The former (e.g., change of face shape) is applied to the whole video sequence directly, while the latter (e.g., change of facial expression or dynamics) is propagated with the guidance of expression or only affects adjacent frames in a given time window. Since users often perform different editing operations in multiple frames, we further present a region-aware fusion approach to fuse diverse editing effects. Our method supports video editing on facial structure and expression movement by sketch, which cannot be achieved by previous works. Both qualitative and quantitative evaluations show the superior editing ability of our system to existing and alternative solutions.
Ask about this paper
Ask your agent about it.
Lune has read the top-tier papers around this one, so every answer names the papers it rests on.
Your agent calls
Lunesearch_papers
Free to start. No credit card required.
Terminal
Install the CLIlune papers get 966d35f4-80df-44c8-9674-e3da06473e90Cited by top-tier papers9
- StyleGANEX: StyleGAN-Based Manipulation Beyond Cropped Aligned FacesShuai Yang, Liming Jiang, Ziwei Liu, Chen Change LoyICCV 2023 · 33 citations
- StreamME: Simplify 3D Gaussian Avatar within Live StreamLuchuan Song, Yang Zhou, Zhan Xu, Yi Zhou et al.SIGGRAPH 2025 · 2 citations
- Sketch3DVE: Sketch-based 3D-Aware Scene Video EditingFeng-Lin Liu, Shi-Yang Li, Yan-Pei Cao, Hongbo Fu et al.SIGGRAPH 2025 · 2 citations
- Bring Clipart to LifeNanxuan Zhao, Shengqi Dang, Hexun Lin, Yang Shi et al.ICCV 2023 · 1 citation
- What Can Human Sketches Do for Object Detection?Pinaki Nath Chowdhury, Ayan Kumar Bhunia, Aneeshan Sain, Subhadeep Koley et al.CVPR 2023
Related papers
- A Latent Transformer for Disentangled Face Editing in Images and VideosXu Yao, Alasdair Newson, Yann Gousseau, Pierre HellierICCV 2021 · 97 citations
- Unsupervised Facial Performance Editing via Vector-Quantized StyleGAN RepresentationsBerkay Kicanaoglu, Pablo Garrido, Gaurav BharajICCV 2023 · 2 citations
- VidStyleODE: Disentangled Video Editing via StyleGAN and NeuralODEsMoayed Haji Ali, Andrew Bond, Levent Karacan, Tolga Birdal et al.ICCV 2023 · 3 citations
- StyleAvatar: Real-time Photo-realistic Portrait Avatar from a Single VideoLizhen Wang, Xiaochen Zhao, Jingxiang Sun, Yuxiang Zhang et al.SIGGRAPH 2023 · 49 citations
- Image2StyleGAN: How to Embed Images Into the StyleGAN Latent Space?Rameen Abdal, Yipeng Qin, Peter WonkaICCV 2019 · 1,195 citations
