CNN-LSTM and Transfer Learning Models for Malware Classification based on Opcodes and API Calls

May 04, 2024 · Declared Dead · 🏛 Knowledge-Based Systems

"No code URL or promise found in abstract"

Evidence collected by the PWNC Scanner

Authors Ahmed Bensaoud, Jugal Kalita arXiv ID 2405.02548 Category cs.CR: Cryptography & Security Cross-listed cs.AI, cs.LG Citations 39 Venue Knowledge-Based Systems Last Checked 4 months ago

Abstract

In this paper, we propose a novel model for a malware classification system based on Application Programming Interface (API) calls and opcodes, to improve classification accuracy. This system uses a novel design of combined Convolutional Neural Network and Long Short-Term Memory. We extract opcode sequences and API Calls from Windows malware samples for classification. We transform these features into N-grams (N = 2, 3, and 10)-gram sequences. Our experiments on a dataset of 9,749,57 samples produce high accuracy of 99.91% using the 8-gram sequences. Our method significantly improves the malware classification performance when using a wide range of recent deep learning architectures, leading to state-of-the-art performance. In particular, we experiment with ConvNeXt-T, ConvNeXt-S, RegNetY-4GF, RegNetY-8GF, RegNetY-12GF, EfficientNetV2, Sequencer2D-L, Swin-T, ViT-G/14, ViT-Ti, ViT-S, VIT-B, VIT-L, and MaxViT-B. Among these architectures, Swin-T and Sequencer2D-L architectures achieved high accuracies of 99.82% and 99.70%, respectively, comparable to our CNN-LSTM architecture although not surpassing it.

📄 View on arXiv 🌐 View on ar5iv 📑 PDF 🎉 Report Code Found

Community Contributions

Found the code? Know the venue? Think something is wrong? Let us know!

📜 Similar Papers

In the same crypt — Cryptography & Security

R.I.P. 👻 Ghosted

Membership Inference Attacks against Machine Learning Models

Reza Shokri, Marco Stronati, ... (+2 more)

cs.CR 🏛 IEEE S&P 📚 4.9K cites 9 years ago

R.I.P. 👻 Ghosted

The Limitations of Deep Learning in Adversarial Settings

Nicolas Papernot, Patrick McDaniel, ... (+4 more)

cs.CR 🏛 IEEE S&P 📚 4.2K cites 10 years ago

R.I.P. 👻 Ghosted

Distillation as a Defense to Adversarial Perturbations against Deep Neural Networks

Nicolas Papernot, Patrick McDaniel, ... (+3 more)

cs.CR 🏛 IEEE S&P 📚 3.2K cites 10 years ago

R.I.P. 👻 Ghosted

Spectre Attacks: Exploiting Speculative Execution

Paul Kocher, Daniel Genkin, ... (+8 more)

cs.CR 🏛 IEEE S&P 📚 2.4K cites 8 years ago

R.I.P. 👻 Ghosted

How To Backdoor Federated Learning

Eugene Bagdasaryan, Andreas Veit, ... (+3 more)

cs.CR 🏛 AISTATS 📚 2.4K cites 7 years ago

R.I.P. 👻 Ghosted

Evasion Attacks against Machine Learning at Test Time

Battista Biggio, Igino Corona, ... (+6 more)

cs.CR 🏛 ECML/PKDD 📚 2.3K cites 8 years ago

Died the same way — 👻 Ghosted

R.I.P. 👻 Ghosted

Federated Learning: Strategies for Improving Communication Efficiency

Jakub Konečný, H. Brendan McMahan, ... (+4 more)

cs.LG 🏛 arXiv 📚 5.2K cites 9 years ago

R.I.P. 👻 Ghosted

In-Datacenter Performance Analysis of a Tensor Processing Unit

Norman P. Jouppi, Cliff Young, ... (+73 more)

cs.AR 🏛 ISCA 📚 5.1K cites 9 years ago

R.I.P. 👻 Ghosted

Deep Convolutional Neural Networks for Computer-Aided Detection: CNN Architectures, Dataset Characteristics and Transfer Learning

Hoo-Chang Shin, Holger R. Roth, ... (+7 more)

cs.CV 🏛 IEEE TMI 📚 4.9K cites 10 years ago

R.I.P. 👻 Ghosted

Explanation in Artificial Intelligence: Insights from the Social Sciences

Tim Miller

cs.AI 🏛 AI 📚 4.9K cites 9 years ago