GeNIS: A modular dataset for network intrusion detection and classification

被引：0

作者：

Silva, Miguel ^{[1
]}

Pinto, Daniela ^{[1
]}

Vitorino, Joao ^{[1
]}

Goncalves, Jose ^{[1
]}

Maia, Eva ^{[1
]}

Praca, Isabel ^{[1
]}

机构：

[1] Polytech Porto ISEP IPP, Sch Engn, Res Grp Intelligent Engn & Comp Adv Innovat & Dev, P-4249015 Porto, Portugal

来源：

DATA IN BRIEF | 2025年 / 60卷

关键词：

Network flow; Packet capture; Attack classification; Anomaly detection; Machine learning; Cybersecurity; Dataset;

D O I：

10.1016/j.dib.2025.111487

中图分类号：

O [数理科学和化学]; P [天文学、地球科学]; Q [生物科学]; N [自然科学总论];

学科分类号：

07 ; 0710 ; 09 ;

摘要：

The development of artificial intelligence solutions for cyberattack detection and classification require high-quality and representative data. However, there is a scarcity of labelled datasets focused on the cyberattacks that target vulnerable small and medium-sized enterprises. To allow organizations to improve their intrusion detection systems according to their types of users, their active services, and the network protocols they use, it is necessary to provide reliable captures of different types of benign and malicious traffic. The GECAD Network Intrusion Scenarios (GeNIS) dataset contains multiple sequential attack scenarios and different types of realistic normal network activity, recorded during advanced network simulations on the Airbus CyberRange platform. The raw network packets were analyzed to generate labelled network flows, with the computation of statistical features to represent the traffic patterns of local and remote attackers, normal users and administrators, and background traffic of an enterprise computer network. GeNIS follows a modular design, providing raw packet capture next generation (PCAPNG) files with over 37 million packets of each intermediate attack step to enable an in-depth analysis with different flow exporters, feature extraction, and feature selection tools, as well as filtered CSV files with over 2.8 million flows created with 5, 10, 30, and 60 s flow intervals. The flows were preprocessed to provide a reliable benchmark dataset with the most relevant features for the training, validation, and testing of robust machine learning and deep learning models.

引用

页数：14

共 50 条

[31] Hierarchical Autoencoder for Network Intrusion Detection
Kye, Hyoseon
Kim, Miru
Kwon, Minhae
[J]. IEEE INTERNATIONAL CONFERENCE ON COMMUNICATIONS (ICC 2022), 2022, : 2700 - 2705
[32] Comparison of Advanced Classification Algorithms Based Intrusion Detection from Real-Time Dataset
R. Aswanandini
C. Deepa
[J]. Automatic Control and Computer Sciences, 2023, 57 : 287 - 295
[33] Comparison of Advanced Classification Algorithms Based Intrusion Detection from Real-Time Dataset
Aswanandini, R.
Deepa, C.
[J]. AUTOMATIC CONTROL AND COMPUTER SCIENCES, 2023, 57 (03) : 287 - 295
[34] Unknown, Atypical and Polymorphic Network Intrusion Detection: A Systematic Survey
Sabeel, Ulya
Heydari, Shahram Shah
El-Khatib, Khalil
Elgazzar, Khalid
[J]. IEEE TRANSACTIONS ON NETWORK AND SERVICE MANAGEMENT, 2024, 21 (01): : 1190 - 1212
[35] Empirical study on multiclass classification-based network intrusion detection
Elmasry, Wisam
Akbulut, Akhan
Zaim, Abdul Halim
[J]. COMPUTATIONAL INTELLIGENCE, 2019, 35 (04) : 919 - 954
[36] On the use of Machine Learning Approaches for the Early Classification in Network Intrusion Detection
Guarino, Idio
Bovenzi, Giampaolo
Di Monda, Davide
Aceto, Giuseppe
Ciuonzo, Domenico
Pescap, Antonio
[J]. 2022 IEEE INTERNATIONAL SYMPOSIUM ON MEASUREMENTS & NETWORKING (M&N 2022), 2022,
[37] Dynamic Deep Forest: An Ensemble Classification Method for Network Intrusion Detection
Hu, Bo
Wang, Jinxi
Zhu, Yifan
Yang, Tan
[J]. ELECTRONICS, 2019, 8 (09)
[38] FEED-FORWARD INTRUSION DETECTION AND CLASSIFICATION ON A SMART GRID NETWORK
Aribisala, Adedayo
Khan, Mohammad S.
Husari, Ghaith
[J]. 2022 IEEE 12TH ANNUAL COMPUTING AND COMMUNICATION WORKSHOP AND CONFERENCE (CCWC), 2022, : 99 - 105
[39] Pelican: A Deep Residual Network for Network Intrusion Detection
Wu, Peilun
Guo, Hui
Moustafa, Nour
[J]. 50TH ANNUAL IEEE/IFIP INTERNATIONAL CONFERENCE ON DEPENDABLE SYSTEMS AND NETWORKS WORKSHOPS (DSN-W 2020), 2020, : 55 - 62
[40] Feature analysis, evaluation and comparisons of classification algorithms based on noisy intrusion dataset
Hussain, Jamal
Lalmuanawma, Samuel
[J]. 2ND INTERNATIONAL CONFERENCE ON INTELLIGENT COMPUTING, COMMUNICATION & CONVERGENCE, ICCC 2016, 2016, 92 : 188 - 198

← 1 2 3 4 5 →