Canonswap: High-Fidelity and Consistent Video Face Swapping Via Canonical Space Modulation
Xiangyang Luo, Ye Zhu, Yunfei Liu, Lijian Lin, Cong Wan, Zijian Cai, Yu Li, Shao-Lun Huang
Abstract
Video face swapping aims to address two primary challenges: effectively transferring the source identity to the target video and accurately preserving the dynamic attributes of the target face, such as head poses, facial expressions, lip-sync, etc. Existing methods mainly focus on achieving high-quality identity transfer but often fall short in maintaining the dynamic attributes of the target face, leading to inconsistent results. We attribute this issue to the inherent coupling of facial appearance and motion in videos. To address this, we propose CanonSwap, a novel video faceswapping framework that decouples motion information from appearance information. Specifically, CanonSwap first eliminates motion-related information, enabling identity modification within a unified canonical space. Subsequently, the swapped feature is reintegrated into the original video space, ensuring the preservation of the target face's dynamic attributes. To further achieve precise identity transfer with minimal artifacts and enhanced realism, we design a Partial Identity Modulation module that adaptively integrates source identity features using a spatial mask to restrict modifications to facial regions. Additionally, we introduce several fine-grained synchronization metrics to comprehensively evaluate the performance of video face swapping methods. Extensive experiments demonstrate that our method significantly outperforms existing approaches in terms of visual quality, temporal consistency, and identity preservation. Our project page are publicly available at https://luoxyhappy.github.io/CanonSwap/.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers4
- DreamID-Omni: Unified Framework for Controllable Human-Centric Audio-Video GenerationXu Guo, Fulong Ye, Qichao Sun, Liyang Chen et al.ICML 2026 · 16 citations
- Beyond the Golden Data: Resolving the Motion-Vision Quality Dilemma via Timestep Selective TrainingXiangyang Luo, Qingyu Li, Yuming Li, Guanbo Huang et al.CVPR 2026 · 3 citations
- FilmWeaver: Weaving Consistent Multi-Shot Videos with Cache-Guided Autoregressive DiffusionXiangyang Luo, Qingyu Li, Xiaokun Liu, Wenyu Qin et al.AAAI 2026 · 3 citations
- Preserving Source Video Realism: High-Fidelity Face Swapping for Cinematic QualityZekai Luo, Zongze Du, Zhouhang Zhu, Hao Zhong et al.CVPR 2026 · 1 citation
Builds on26
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- FaceForensics++: Learning to Detect Manipulated Facial ImagesAndreas Rössler, Davide Cozzolino, Luisa Verdoliva, Christian Riess et al.ICCV 2019 · 2,966 citations
- Video Diffusion ModelsJonathan Ho, Tim Salimans, Alexey A. Gritsenko, William Chan et al.NeurIPS 2022 · 2,948 citations
- AnimateDiff: Animate Your Personalized Text-to-Image Diffusion Models without Specific TuningYuwei Guo, Ceyuan Yang, Anyi Rao, Zhengyang Liang et al.ICLR 2024 · 1,493 citations
- SimSwap: An Efficient Framework For High Fidelity Face SwappingRenwang Chen, Xuanhong Chen, Bingbing Ni, Yanhao GeACM MM 2020 · 409 citations
Related papers
- DynamicFace: High-Quality and Consistent Face Swapping for Image and Video Using Composable 3D Facial PriorsRunqi Wang, Yang Chen, Sijie Xu, Tianyao He et al.ICCV 2025 · 7 citations
- Controllable and Expressive One-Shot Video Head SwappingChaonan Ji, Jinwei Qi, Peng Zhang, Bang Zhang et al.ICCV 2025
- High-Fidelity Diffusion Face Swapping with ID-Constrained Facial ConditioningDailan He, Xiahong Wang, Shulun Wang, Hao Shao et al.CVPR 2026 · 5 citations
- CodeSwap: Symmetrically Face Swapping Based on Prior CodebookXiangyang Luo, Xin Zhang, Yifan Xie, Xinyi Tong et al.ACM MM 2024 · 9 citations
- High Fidelity Face Swapping via Semantics Disentanglement and Structure EnhancementFengyuan Liu, Lingyun Yu, Hongtao Xie, Chuanbin Liu et al.ACM MM 2023 · 1 citation
