Tackling imbalanced data in cybersecurity with transfer learning: a case with ROP payload detection

Wang, Haizhou; Singhal, Anoop; Liu, Peng

doi:10.1186/s42400-022-00135-8

Citation Details

Tackling imbalanced data in cybersecurity with transfer learning: a case with ROP payload detection

Abstract In recent years, deep learning gained proliferating popularity in the cybersecurity application domain, since when being compared to traditional machine learning methods, it usually involves less human efforts, produces better results, and provides better generalizability. However, the imbalanced data issue is very common in cybersecurity, which can substantially deteriorate the performance of the deep learning models. This paper introduces a transfer learning based method to tackle the imbalanced data issue in cybersecurity using return-oriented programming payload detection as a case study. We achieved 0.0290 average false positive rate, 0.9705 average F1 score and 0.9521 average detection rate on 3 different target domain programs using 2 different source domain programs, with 0 benign training data sample in the target domain. The performance improvement compared to the baseline is a trade-off between false positive rate and detection rate. Using our approach, the total number of false positives is reduced by 23.16%, and as a trade-off, the number of detected malicious samples decreases by 0.68%. more »

Award ID(s):: 2140175 2019340

PAR ID:: 10389769

Author(s) / Creator(s):: Wang, Haizhou; Singhal, Anoop; Liu, Peng

Publisher / Repository:: Springer Science + Business Media

Date Published:: 2023-01-05

Journal Name:: Cybersecurity

Volume:: 6

Issue:: 1

ISSN:: 2523-3246

Format(s):: Medium: X

Sponsoring Org:: National Science Foundation

Journal Article:
https://doi.org/10.1186/s42400-022-00135-8

More Like this