MalCL: Leveraging GAN-Based Generative Replay to Combat Catastrophic Forgetting in Malware Classification
Jimin Park, AHyun Ji, Minji Park, Mohammad Saidur Rahman, Se Eun Oh
摘要
Continual Learning (CL) for malware classification tackles the rapidly evolving nature of malware threats and the frequent emergence of new types. Generative Replay (GR)based CL systems utilize a generative model to produce synthetic versions of past data, which are then combined with new data to retrain the primary model. Traditional machine learning techniques in this domain often struggle with catastrophic forgetting, where a model's performance on old data degrades over time. In this paper, we introduce a GR-based CL system that employs Generative Adversarial Networks (GANs) with feature matching loss to generate high-quality malware samples. Additionally, we implement innovative selection schemes for replay samples based on the model's hidden representations. Our comprehensive evaluation across Windows and Android malware datasets in a class-incremental learning scenariowhere new classes are introduced continuously over multiple tasks -demonstrates substantial performance improvements over previous methods. For example, our system achieves an average accuracy of 55% on Windows malware samples, significantly outperforming other GR-based models by 28%. This study provides practical insights for advancing GR-based malware classification systems. The implementation is available at https://github.com/MalwareReplayGAN/ MalCL 1 .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- LAMDA: A Longitudinal Android Malware Benchmark for Concept Drift AnalysisMd Ahsanul Haque, Ismail Hossain, Md Mahmuduzzaman Kamol, Md Jahangir Alam 等ICLR 2026 · 被引用 14 次
- HyperGLLM: An Efficient Framework for Endpoint Threat Detection via Hypergraph-Enhanced Large Language ModelsHongyi Zhou, Jianfeng Pan, Min Peng, Shaomang Huang 等AAAI 2026
它引用的顶会 Paper6
- What is being transferred in transfer learning?Behnam Neyshabur, Hanie Sedghi, Chiyuan ZhangNeurIPS 2020 · 被引用 654 次
- DDGR: Continual Learning with Deep Diffusion-based Generative ReplayRui Gao, Weiwei LiuICML 2023 · 被引用 101 次
- Classifying Sequences of Extreme Length with Constant Memory Applied to Malware DetectionEdward Raff, William Fleshman, Richard Zak, Hyrum S. Anderson 等AAAI 2021 · 被引用 70 次
- Augmented Memory Replay-based Continual Learning Approaches for Network Intrusion DetectionSuresh Kumar Amalapuram, Sumohana S. Channappayya, Bheemarjuna Reddy TammaNeurIPS 2023 · 被引用 42 次
- Continuous Learning for Android Malware DetectionYizheng Chen, Zhoujie Ding, David A. WagnerUSENIX Security 2023
相关 Paper
- Guided Retraining to Enhance the Detection of Difficult Android MalwareNadia Daoudi, Kevin Allix, Tegawendé F. Bissyandé, Jacques KleinISSTA 2023 · 被引用 4 次
- Semantics-Driven Generative Replay for Few-Shot Class Incremental LearningAishwarya Agarwal, Biplab Banerjee, Fabio Cuzzolin, Subhasis ChaudhuriACM MM 2022 · 被引用 26 次
- GCR: Gradient Coreset based Replay Buffer Selection for Continual LearningRishabh Tiwari, KrishnaTeja Killamsetty, Rishabh K. Iyer, Pradeep ShenoyCVPR 2022 · 被引用 102 次
- RECALL: Replay-based Continual Learning in Semantic SegmentationAndrea Maracani, Umberto Michieli, Marco Toldo, Pietro ZanuttighICCV 2021 · 被引用 148 次
- DeepCollaboration: Collaborative Generative and Discriminative Models for Class Incremental LearningBo Cui, Guyue Hu, Shan YuAAAI 2021 · 被引用 11 次
