Classifying Sequences of Extreme Length with Constant Memory Applied to Malware Detection
Edward Raff, William Fleshman, Richard Zak, Hyrum S. Anderson, Bobby Filar, Mark McLean
摘要
Recent works within machine learning have been tackling inputs of ever-increasing size, with cybersecurity presenting sequence classification problems of particularly extreme lengths. In the case of Windows executable malware detection, inputs may exceed 100 MB, which corresponds to a time series with T = 100, 000, 000 steps. To date, the closest approach to handling such a task is MalConv, a convolutional neural network capable of processing up to T = 2, 000, 000 steps. The O(T ) memory of CNNs has prevented further application of CNNs to malware. In this work, we develop a new approach to temporal max pooling that makes the required memory invariant to the sequence length T . This makes MalConv 116× more memory efficient, and up to 25.8× faster to train on its original dataset, while removing the input length restrictions to MalConv. We re-invest these gains into improving the Mal-Conv architecture by developing a new Global Channel Gating design, giving us an attention mechanism capable of learning feature interactions across 100 million time steps in an efficient manner, a capability lacked by the original MalConv CNN. Our implementation can be found at https://github.com/ NeuromorphicComputationResearchProgram/MalConv2
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper7
- RS-Del: Edit Distance Robustness Certificates for Sequence Classifiers via Randomized DeletionZhuoqun Huang, Neil G. Marchant, Keane Lucas, Lujo Bauer 等NeurIPS 2023 · 被引用 24 次
- Recasting Self-Attention with Holographic Reduced RepresentationsMohammad Mahmudul Alam, Edward Raff, Stella Biderman, Tim Oates 等ICML 2023 · 被引用 18 次
- MalCL: Leveraging GAN-Based Generative Replay to Combat Catastrophic Forgetting in Malware ClassificationJimin Park, AHyun Ji, Minji Park, Mohammad Saidur Rahman 等AAAI 2025 · 被引用 12 次
- Beyond Raw Bytes: Towards Large Malware Language ModelsLuke Kurlandski, Harel Berger, Yin Pan, Matthew WrightNDSS 2026 · 被引用 5 次
- MPass: Bypassing Learning-based Static Malware DetectorsJialai Wang, Wenjie Qu, Yi Rong, Han Qiu 等DAC 2023 · 被引用 4 次
它引用的顶会 Paper3
- Reformer: The Efficient TransformerNikita Kitaev, Lukasz Kaiser, Anselm LevskayaICLR 2020 · 被引用 2,878 次
- A New Burrows Wheeler Transform Markov DistanceEdward Raff, Charles Nicholas, Mark McLeanAAAI 2020 · 被引用 13 次
- An Observational Investigation of Reverse Engineers' ProcessesDaniel Votipka, Seth M. Rabin, Kristopher K. Micinski, Jeffrey S. Foster 等USENIX Security 2020
相关 Paper
- Adversarial Training for Raw-Binary Malware ClassifiersKeane Lucas, Samruddhi Pai, Weiran Lin, Lujo Bauer 等USENIX Security 2023
- Dynamic Malware Analysis with Feature Engineering and Feature LearningZhaoqi Zhang, Panpan Qi, Wei WangAAAI 2020 · 被引用 153 次
- MalDetectFormer: Leveraging Sparse SpatioTemporal Information for Effective Malicious Traffic DetectionShuai Zhang, Yu Fan, Haoyi Zhou, Bo LiAAAI 2025 · 被引用 1 次
- Time-aware Large Kernel ConvolutionsVasileios Lioutas, Yuhong GuoICML 2020 · 被引用 30 次
- Sequence Modeling with Multiresolution Convolutional MemoryJiaxin Shi, Ke Alexander Wang, Emily B. FoxICML 2023 · 被引用 24 次
