Attribute-Preserving Pseudo-Labeling for Diffusion-Based Face Swapping
Jiwon Kang, Yeji Choi, JoungBin Lee, Wooseok Jang, Jinhyeok Choi, Taekeun Kang, Yongjae Park, Myungin Kim, Seungryong Kim
摘要
Face swapping aims to transfer the identity of a source face onto a target face while preserving target-specific attributes such as pose, expression, lighting, skin tone, and makeup. However, since real face-swapping ground truth is unavailable, achieving both accurate identity transfer and high-quality attribute preservation remains challenging. Although recent diffusion-based approaches attempt to improve visual fidelity through conditional inpainting on masked target images, the masked condition removes crucial appearance cues, resulting in plausible yet misaligned attributes due to the lack of explicit supervision. To address these limitations, we propose APPLE (Attribute-Preserving Pseudo-Labeling for Diffusion-Based Face Swapping) a diffusion-based teacher–student framework that enhances attribute fidelity through attribute-aware pseudo-label supervision. First, we reformulate face swapping as a conditional deblurring task to more faithfully preserve target-specific attributes such as lighting, skin tone, and makeup. In addition, we introduce an attribute-aware inversion scheme to further improve detailed attribute preservation. Through an elaborate attribute-preserving design for teacher learning, APPLE produces high-quality pseudo triplets that explicitly provide the student with direct face-swapping supervision. Overall, APPLE achieves state-of-the-art performance in terms of attribute preservation and identity transfer, producing more photorealistic and target-faithful results. Code will be made publicly available.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper26
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- LoRA: Low-Rank Adaptation of Large Language ModelsEdward J. Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu 等ICLR 2022 · 被引用 18,833 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- Denoising Diffusion Implicit ModelsJiaming Song, Chenlin Meng, Stefano ErmonICLR 2021 · 被引用 11,743 次
相关 Paper
- High-Fidelity Diffusion Face Swapping with ID-Constrained Facial ConditioningDailan He, Xiahong Wang, Shulun Wang, Hao Shao 等CVPR 2026 · 被引用 5 次
- FaceController: Controllable Attribute Editing for Face in the WildZhiliang Xu, Xiyu Yu, Zhibin Hong, Zhen Zhu 等AAAI 2021 · 被引用 49 次
- DynamicFace: High-Quality and Consistent Face Swapping for Image and Video Using Composable 3D Facial PriorsRunqi Wang, Yang Chen, Sijie Xu, Tianyao He 等ICCV 2025 · 被引用 7 次
- High Fidelity Face Swapping via Semantics Disentanglement and Structure EnhancementFengyuan Liu, Lingyun Yu, Hongtao Xie, Chuanbin Liu 等ACM MM 2023 · 被引用 1 次
- One-stage Context and Identity Hallucination NetworkYinglu Liu, Mingcan Xiang, Hailin Shi, Tao MeiACM MM 2021 · 被引用 2 次
