Do You Remember? Overcoming Catastrophic Forgetting for Fake Audio Detection
Xiaohui Zhang, Jiangyan Yi, Jianhua Tao, Chenglong Wang, Chu Yuan Zhang
摘要
Current fake audio detection algorithms have achieved promising performances on most datasets. However, their performance may be significantly degraded when dealing with audio of a different dataset. The orthogonal weight modification to overcome catastrophic forgetting does not consider the similarity of genuine audio across different datasets. To overcome this limitation, we propose a continual learning algorithm for fake audio detection to overcome catastrophic forgetting, called Regularized Adaptive Weight Modification (RAWM). When fine-tuning a detection network, our approach adaptively computes the direction of weight modification according to the ratio of genuine utterances and fake utterances. The adaptive modification direction ensures the network can effectively detect fake audio on the new dataset while preserving its knowledge of old model, thus mitigating catastrophic forgetting. In addition, genuine audio collected from quite different acoustic conditions may skew their feature distribution, so we introduce a regularization constraint to force the network to remember the old distribution in this regard. Our method can easily be generalized to related fields, like speech emotion recognition. We also evaluate our approach across multiple datasets and obtain a significant performance improvement on cross-dataset experiments.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- Improving Generalization for AI-Synthesized Voice DetectionHainan Ren, Li Lin, Chun-Hao Liu, Xin Wang 等AAAI 2025 · 被引用 13 次
- Continual Audio-Visual Sound SeparationWeiguo Pian, Yiyang Nan, Shijian Deng, Shentong Mo 等NeurIPS 2024 · 被引用 11 次
- Region-Based Optimization in Continual Learning for Audio Deepfake DetectionYujie Chen, Jiangyan Yi, Cunhang Fan, Jianhua Tao 等AAAI 2025 · 被引用 10 次
- MusicDET: Zero-Shot AI-Generated Music DetectionChaolei Han, Hongsong Wang, Jie GuiICML 2026
- Multimodal Representation Learning by Alternating Unimodal AdaptationXiaohui Zhang, Jaehong Yoon, Mohit Bansal, Huaxiu YaoCVPR 2024
它引用的顶会 Paper4
- wav2vec 2.0: A Framework for Self-Supervised Learning of Speech RepresentationsAlexei Baevski, Yuhao Zhou, Abdelrahman Mohamed, Michael AuliNeurIPS 2020 · 被引用 9,451 次
- Overcoming Catastrophic Forgetting With Unlabeled Data in the WildKibok Lee, Kimin Lee, Jinwoo Shin, Honglak LeeICCV 2019 · 被引用 231 次
- Incremental Few-Shot Object DetectionJuan-Manuel Pérez-Rúa, Xiatian Zhu, Timothy M. Hospedales, Tao XiangCVPR 2020
- Modeling the Background for Incremental Learning in Semantic SegmentationFabio Cermelli, Massimiliano Mancini, Samuel Rota Bulò, Elisa Ricci 等CVPR 2020
相关 Paper
- What to Remember: Self-Adaptive Continual Learning for Audio Deepfake DetectionXiaohui Zhang, Jiangyan Yi, Chenglong Wang, Chu Yuan Zhang 等AAAI 2024 · 被引用 44 次
- Choose Your Expert: Uncertainty-Guided Expert Selection for Continual Deepfake DetectionXueyi Zhang, Peiyin Zhu, Jinping Sui, Xiaoda Yang 等ACM MM 2025
- DevFD : Developmental Face Forgery Detection by Learning Shared and Orthogonal LoRA SubspacesTianshuo Zhang, Li Gao, Siran Peng, Xiangyu Zhu 等NeurIPS 2025 · 被引用 4 次
- Exploring The Forgetting in Adversarial Training: A Novel Method for Enhancing RobustnessXianglu Wang, Hu DingICLR 2025
- CoReD: Generalizing Fake Media Detection with Continual Representation using DistillationMinha Kim, Shahroz Tariq, Simon S. WooACM MM 2021 · 被引用 48 次
