Do You Remember? Overcoming Catastrophic Forgetting for Fake Audio Detection
Xiaohui Zhang, Jiangyan Yi, Jianhua Tao, Chenglong Wang, Chu Yuan Zhang
Abstract
Current fake audio detection algorithms have achieved promising performances on most datasets. However, their performance may be significantly degraded when dealing with audio of a different dataset. The orthogonal weight modification to overcome catastrophic forgetting does not consider the similarity of genuine audio across different datasets. To overcome this limitation, we propose a continual learning algorithm for fake audio detection to overcome catastrophic forgetting, called Regularized Adaptive Weight Modification (RAWM). When fine-tuning a detection network, our approach adaptively computes the direction of weight modification according to the ratio of genuine utterances and fake utterances. The adaptive modification direction ensures the network can effectively detect fake audio on the new dataset while preserving its knowledge of old model, thus mitigating catastrophic forgetting. In addition, genuine audio collected from quite different acoustic conditions may skew their feature distribution, so we introduce a regularization constraint to force the network to remember the old distribution in this regard. Our method can easily be generalized to related fields, like speech emotion recognition. We also evaluate our approach across multiple datasets and obtain a significant performance improvement on cross-dataset experiments.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers5
- Improving Generalization for AI-Synthesized Voice DetectionHainan Ren, Li Lin, Chun-Hao Liu, Xin Wang et al.AAAI 2025 · 13 citations
- Continual Audio-Visual Sound SeparationWeiguo Pian, Yiyang Nan, Shijian Deng, Shentong Mo et al.NeurIPS 2024 · 11 citations
- Region-Based Optimization in Continual Learning for Audio Deepfake DetectionYujie Chen, Jiangyan Yi, Cunhang Fan, Jianhua Tao et al.AAAI 2025 · 10 citations
- MusicDET: Zero-Shot AI-Generated Music DetectionChaolei Han, Hongsong Wang, Jie GuiICML 2026
- Multimodal Representation Learning by Alternating Unimodal AdaptationXiaohui Zhang, Jaehong Yoon, Mohit Bansal, Huaxiu YaoCVPR 2024
Builds on4
- wav2vec 2.0: A Framework for Self-Supervised Learning of Speech RepresentationsAlexei Baevski, Yuhao Zhou, Abdelrahman Mohamed, Michael AuliNeurIPS 2020 · 9,451 citations
- Overcoming Catastrophic Forgetting With Unlabeled Data in the WildKibok Lee, Kimin Lee, Jinwoo Shin, Honglak LeeICCV 2019 · 231 citations
- Incremental Few-Shot Object DetectionJuan-Manuel Pérez-Rúa, Xiatian Zhu, Timothy M. Hospedales, Tao XiangCVPR 2020
- Modeling the Background for Incremental Learning in Semantic SegmentationFabio Cermelli, Massimiliano Mancini, Samuel Rota Bulò, Elisa Ricci et al.CVPR 2020
Related papers
- What to Remember: Self-Adaptive Continual Learning for Audio Deepfake DetectionXiaohui Zhang, Jiangyan Yi, Chenglong Wang, Chu Yuan Zhang et al.AAAI 2024 · 44 citations
- Choose Your Expert: Uncertainty-Guided Expert Selection for Continual Deepfake DetectionXueyi Zhang, Peiyin Zhu, Jinping Sui, Xiaoda Yang et al.ACM MM 2025
- DevFD : Developmental Face Forgery Detection by Learning Shared and Orthogonal LoRA SubspacesTianshuo Zhang, Li Gao, Siran Peng, Xiangyu Zhu et al.NeurIPS 2025 · 4 citations
- Exploring The Forgetting in Adversarial Training: A Novel Method for Enhancing RobustnessXianglu Wang, Hu DingICLR 2025
- CoReD: Generalizing Fake Media Detection with Continual Representation using DistillationMinha Kim, Shahroz Tariq, Simon S. WooACM MM 2021 · 48 citations
