Region-Based Optimization in Continual Learning for Audio Deepfake Detection
Yujie Chen, Jiangyan Yi, Cunhang Fan, Jianhua Tao, Yong Ren, Siding Zeng, Chu Yuan Zhang, Xinrui Yan, Hao Gu, Jun Xue, Chenglong Wang, Zhao Lv, Xiaohui Zhang
摘要
Rapid advancements in speech synthesis and voice conversion bring convenience but also new security risks, creating an urgent need for effective audio deepfake detection. Although current models perform well, their effectiveness diminishes when confronted with the diverse and evolving nature of real-world deepfakes. To address this issue, we propose a continual learning method named Region-Based Optimization (RegO) for audio deepfake detection. Specifically, we use the Fisher information matrix to measure important neuron regions for real and fake audio detection, dividing them into four regions. First, we directly fine-tune the less important regions to quickly adapt to new tasks. Next, we apply gradient optimization in parallel for regions important only to real audio detection, and in orthogonal directions for regions important only to fake audio detection. For regions that are important to both, we use sample proportion-based adaptive gradient optimization. This region-adaptive optimization ensures an appropriate trade-off between memory stability and learning plasticity. Additionally, to address the increase of redundant neurons from old tasks, we further introduce the Ebbinghaus forgetting mechanism to release them, thereby promoting the model's ability to learn more generalized discriminative features. Experimental results show our method achieves a 21.3% improvement in EER over the state-of-theart continual learning approach RWM for audio deepfake detection. Moreover, the effectiveness of RegO extends beyond the audio deepfake detection domain, showing potential significance in other tasks, such as image recognition. The code is available at https://github.com/cyjie429/RegO
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper7
- Addressing Loss of Plasticity and Catastrophic Forgetting in Continual LearningMohamed Elsayed, A. Rupam MahmoodICLR 2024 · 被引用 52 次
- Prompt Gradient Projection for Continual LearningJingyang Qiao, Zhizhong Zhang, Xin Tan, Chengwei Chen 等ICLR 2024 · 被引用 47 次
- What to Remember: Self-Adaptive Continual Learning for Audio Deepfake DetectionXiaohui Zhang, Jiangyan Yi, Chenglong Wang, Chu Yuan Zhang 等AAAI 2024 · 被引用 44 次
- Do You Remember? Overcoming Catastrophic Forgetting for Fake Audio DetectionXiaohui Zhang, Jiangyan Yi, Jianhua Tao, Chenglong Wang 等ICML 2023 · 被引用 33 次
- Progressive Prompts: Continual Learning for Language ModelsAnastasia Razdaibiedina, Yuning Mao, Rui Hou, Madian Khabsa 等ICLR 2023 · 被引用 15 次
相关 Paper
- CoReD: Generalizing Fake Media Detection with Continual Representation using DistillationMinha Kim, Shahroz Tariq, Simon S. WooACM MM 2021 · 被引用 48 次
- Choose Your Expert: Uncertainty-Guided Expert Selection for Continual Deepfake DetectionXueyi Zhang, Peiyin Zhu, Jinping Sui, Xiaoda Yang 等ACM MM 2025
- AVFF: Audio-Visual Feature Fusion for Video Deepfake DetectionTrevine Oorloff, Surya Koppisetti, Nicolò Bonettini, Divyaraj Solanki 等CVPR 2024 · 被引用 51 次
- Joint Audio-Visual Deepfake DetectionYipin Zhou, Ser-Nam LimICCV 2021 · 被引用 232 次
- Audio Deepfake Detection with Self-Supervised XLS-R and SLS ClassifierQishan Zhang, Shuangbing Wen, Tao HuACM MM 2024 · 被引用 54 次
