Low-sample classification in NIDS using the EC-GAN

Zekan, Marko; Tomičić, Igor; Schatten, Markus

izvor podataka: crosbi !

Low-sample classification in NIDS using the EC-GAN (CROSBI ID 319286)

Prilog u časopisu | izvorni znanstveni rad | međunarodna recenzija

Zekan, Marko ; Tomičić, Igor ; Schatten, Markus Low-sample classification in NIDS using the EC-GAN // Journal of universal computer science, 28 (2022), 12; 1330-1346. doi: 10.3897/jucs.85703

Podaci o odgovornosti

Autori

Zekan, Marko ; Tomičić, Igor ; Schatten, Markus

Osnovni podaci na izvornom jeziku
Osnovni podaci na ostalim jezicima

Jezik

engleski

Naslov

Low-sample classification in NIDS using the EC-GAN

Sažetak

Numerous advanced methods have been applied throughout the years for the use in Network Intrusion Detection Systems (NIDS). Among these are various Deep Learning models, which have shown great success for attack classification. Nevertheless, false positive rate and detection rate of these systems remains a concern. This is mostly because of the low-sample, imbalanced nature of realistic datasets, which make models challenging to train. Considering this, we applied a novel semi-supervised EC-GAN method for network flow classifi- cation of CIC-IDS-2017 dataset. EC-GAN uses synthetic data to aid the training of a supervised classifier on low- sample data. To achieve this, we modified the original EC-GAN to work with tabular data. In our approach, WCGAN-GP is used for synthetic tabular data generation, while a simple deep neural network is used for classification. The conditional nature of WCGAN-GP diminishes the class imbalance problem, while GAN itself solves the low-sample problem. This approach was successful in generating believable synthetic data, which was consequently used for training and testing the EC-GAN. To obtain our results, we trained a classifier on progressively smaller versions of the CIC-DIS-2017 dataset, first via a novel EC-GAN method and then in the conventional way, without the help of synthetic data. We then compared these two sets of results with another author’s results using accuracy, false positive rate, detection rate and macro F1 score as metrics. Our results showed that supervised classifier trained with EC-GAN can achieve significant results even when trained on as little as 25% of the original imbalanced dataset.

Ključne riječi

cybersecurity ; network security ; GAN ; NIDS ; synthetic tabular data ; classification ; semi-supervised learning ; Wasserstein GAN

Napomena

nije evidentirano

Jezik

nije evidentirano

Naslov

nije evidentirano

Sažetak

nije evidentirano

Ključne riječi

nije evidentirano

Napomena

nije evidentirano

Podaci o izdanju

Časopis

Journal of universal computer science

Volumen (broj)

28 (12)

Godina

2022.

Stranice rada

1330-1346

Status objave rada

objavljeno

ISSN

0948-695X

e-ISSN

0948-6968

DOI

10.3897/jucs.85703

Povezanost rada

Povezane osobe

Igor Tomičić (autor/i)

Markus Schatten (autor/i)

Povezane ustanove

Fakultet organizacije i informatike (016) (autorova ustanova)

Povezani projekti

Orkestracija hibridnih metoda umjetne inteligencije za računalne igre (rezultat rada na projektu)

Područje

Informacijske i komunikacijske znanosti, Računarstvo

Poveznice

doi.org

lib.jucs.org

Indeksiranost

Scopus

Current Contents Connect (CCC)

Web of Science Core Collection, Science Citation Index Expanded (WoSCC-SCI-Exp)

Web of Science Core Collection, SCI-Exp, SSCI & A&HCI (WoSCC-SCI-Exp, SSCI, A&HCI)