Preserving Fairness Generalization in Deepfake Detection
Li Lin, Xinan He, Yan Ju, Xin Wang, Feng Ding, Shu Hu
Abstract
Although effective deepfake detection models have been developed in recent years, recent studies have revealed that these models can result in unfair performance disparities among demographic groups, such as race and gender. This can lead to particular groups facing unfair targeting or exclusion from detection, potentially allowing misclassified deepfakes to manipulate public opinion and undermine trust in the model. The existing method for addressing this problem is providing a fair loss function. It shows good fairness performance for intra-domain evaluation but does not maintain fairness for cross-domain testing. This highlights the significance of fairness generalization in the fight against deepfakes. In this work, we propose the first method to address the fairness generalization problem in deepfake detection by simultaneously considering features, loss, and optimization aspects. Our method employs disentanglement learning to extract demographic and domain-agnostic forgery features, fusing them to encourage fair learning across a flattened loss landscape. Extensive experiments on prominent deepfake datasets demonstrate our method's effectiveness, surpassing state-of-the-art approaches in preserving fairness during cross-domain deepfake detection. The code is available at https://github.com/Purdue-M2/Fairness-Generalization .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 1333386b-5ec3-46bb-b32b-dd723971c8aaCited by top-tier papers21
- FreqBlender: Enhancing DeepFake Detection by Blending Frequency KnowledgeHanzhe Li, Jiaran Zhou, Yuezun Li, Baoyuan Wu et al.NeurIPS 2024 · 96 citations
- Improving Generalization for AI-Synthesized Voice DetectionHainan Ren, Li Lin, Chun-Hao Liu, Xin Wang et al.AAAI 2025 · 13 citations
- Fair Deepfake Detectors Can GeneralizeHarry Cheng, Ming-Hui Liu, Yangyang Guo, Tianyi Wang et al.NeurIPS 2025 · 11 citations
- Thinking Racial Bias in Fair Forgery Detection: Models, Datasets and EvaluationsDecheng Liu, Zongqi Wang, Chunlei Peng, Nannan Wang et al.AAAI 2025 · 11 citations
- BlurGuard: A Simple Approach for Robustifying Image Protection Against AI-Powered EditingJinsu Kim, Yunhun Nam, Minseon Kim, Sangpil Kim et al.NeurIPS 2025 · 7 citations
Builds on17
- Diffusion Models Beat GANs on Image SynthesisPrafulla Dhariwal, Alexander Quinn NicholNeurIPS 2021 · 13,211 citations
- FaceForensics++: Learning to Detect Manipulated Facial ImagesAndreas Rössler, Davide Cozzolino, Luisa Verdoliva, Christian Riess et al.ICCV 2019 · 2,966 citations
- Sharpness-aware Minimization for Efficiently Improving GeneralizationPierre Foret, Ariel Kleiner, Hossein Mobahi, Behnam NeyshaburICLR 2021 · 1,861 citations
- NVAE: A Deep Hierarchical Variational AutoencoderArash Vahdat, Jan KautzNeurIPS 2020 · 1,141 citations
- Symmetric Cross Entropy for Robust Learning With Noisy LabelsYisen Wang, Xingjun Ma, Zaiyi Chen, Yuan Luo et al.ICCV 2019 · 1,125 citations
Related papers
- Decoupling Bias, Aligning Distributions: Synergistic Fairness Optimization for Deepfake DetectionFeng Ding, Wenhui Yi, Yunpeng Zhou, Xinan He et al.CVPR 2026 · 3 citations
- Open-Unfairness Adversarial Mitigation for Generalized Deepfake DetectionZhaoyang Li, Zhu Teng, Baopeng Zhang, Jianping FanICCV 2025 · 1 citation
- Rethinking Individual Fairness in Deepfake DetectionAryana Hou, Li Lin, Justin Li, Shu HuACM MM 2025 · 1 citation
- UCF: Uncovering Common Features for Generalizable Deepfake DetectionZhiyuan Yan, Yong Zhang, Yanbo Fan, Baoyuan WuICCV 2023 · 264 citations
- DiffusionFake: Enhancing Generalization in Deepfake Detection via Guided Stable DiffusionKe Sun, Shen Chen, Taiping Yao, Hong Liu et al.NeurIPS 2024 · 57 citations
