DeepFaceVideoEditing: sketch-based deep editing of face videos
Feng-Lin Liu, Shu-Yu Chen, Yu-Kun Lai, Chunpeng Li, Yue-Ren Jiang, Hongbo Fu, Lin Gao
摘要
Sketches, which are simple and concise, have been used in recent deep image synthesis methods to allow intuitive generation and editing of facial images. However, it is nontrivial to extend such methods to video editing due to various challenges, ranging from appropriate manipulation propagation and fusion of multiple editing operations to ensure temporal coherence and visual quality. To address these issues, we propose a novel sketch-based facial video editing framework, in which we represent editing manipulations in latent space and propose specific propagation and fusion modules to generate high-quality video editing results based on StyleGAN3. Specifically, we first design an optimization approach to represent sketch editing manipulations by editing vectors, which are propagated to the whole video sequence using a proper strategy to cope with different editing needs. Specifically, input editing operations are classified into two categories: temporally consistent editing and temporally variant editing. The former (e.g., change of face shape) is applied to the whole video sequence directly, while the latter (e.g., change of facial expression or dynamics) is propagated with the guidance of expression or only affects adjacent frames in a given time window. Since users often perform different editing operations in multiple frames, we further present a region-aware fusion approach to fuse diverse editing effects. Our method supports video editing on facial structure and expression movement by sketch, which cannot be achieved by previous works. Both qualitative and quantitative evaluations show the superior editing ability of our system to existing and alternative solutions.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
引用它的顶会 Paper9
- StyleGANEX: StyleGAN-Based Manipulation Beyond Cropped Aligned FacesShuai Yang, Liming Jiang, Ziwei Liu, Chen Change LoyICCV 2023 · 被引用 33 次
- StreamME: Simplify 3D Gaussian Avatar within Live StreamLuchuan Song, Yang Zhou, Zhan Xu, Yi Zhou 等SIGGRAPH 2025 · 被引用 2 次
- Sketch3DVE: Sketch-based 3D-Aware Scene Video EditingFeng-Lin Liu, Shi-Yang Li, Yan-Pei Cao, Hongbo Fu 等SIGGRAPH 2025 · 被引用 2 次
- Bring Clipart to LifeNanxuan Zhao, Shengqi Dang, Hexun Lin, Yang Shi 等ICCV 2023 · 被引用 1 次
- What Can Human Sketches Do for Object Detection?Pinaki Nath Chowdhury, Ayan Kumar Bhunia, Aneeshan Sain, Subhadeep Koley 等CVPR 2023
相关 Paper
- A Latent Transformer for Disentangled Face Editing in Images and VideosXu Yao, Alasdair Newson, Yann Gousseau, Pierre HellierICCV 2021 · 被引用 97 次
- Unsupervised Facial Performance Editing via Vector-Quantized StyleGAN RepresentationsBerkay Kicanaoglu, Pablo Garrido, Gaurav BharajICCV 2023 · 被引用 2 次
- VidStyleODE: Disentangled Video Editing via StyleGAN and NeuralODEsMoayed Haji Ali, Andrew Bond, Levent Karacan, Tolga Birdal 等ICCV 2023 · 被引用 3 次
- StyleAvatar: Real-time Photo-realistic Portrait Avatar from a Single VideoLizhen Wang, Xiaochen Zhao, Jingxiang Sun, Yuxiang Zhang 等SIGGRAPH 2023 · 被引用 49 次
- Image2StyleGAN: How to Embed Images Into the StyleGAN Latent Space?Rameen Abdal, Yipeng Qin, Peter WonkaICCV 2019 · 被引用 1,195 次
