| Naslov: | Utilizing large language models for cyber attack data generation : magistrsko delo |
|---|
| Avtorji: | ID Zgaga, Bine (Avtor) ID Beranič, Tina (Mentor) Več o mentorju...  ID Kaymaz, Sevtap Duman (Komentor) |
| Datoteke: | MAG_Zgaga_Bine_2026.pdf (3,40 MB) MD5: 22AB2BECFA652F6BBDBC0C6F41A89C7D
|
|---|
| Jezik: | Angleški jezik |
|---|
| Vrsta gradiva: | Magistrsko delo/naloga |
|---|
| Tipologija: | 2.09 - Magistrsko delo |
|---|
| Organizacija: | FERI - Fakulteta za elektrotehniko, računalništvo in informatiko
|
|---|
| Opis: | This master’s thesis addresses the challenge of limited, imbalanced, and restricted-access datasets in the field of cybersecurity, which significantly hinder the development and evaluation of intrusion detection systems. To overcome these limitations, the thesis explores the use of synthetic data generation as a viable and effective alternative for producing realistic and diverse cyber-attack data while preserving privacy and improving dataset usability.
The work presents a systematic review of existing approaches to synthetic data generation based on Generative Adversarial Networks and Large Language Models, analysing their fundamental characteristics, advantages, and limitations in the context of cyber-attack generation. Particular emphasis is placed on their suitability for creating realistic and diverse attack patterns relevant to intrusion detection research.
The experimental part of the thesis investigates and compares multiple approaches to synthetic data generation, including standalone generative models and a combined methodology. The generated data are evaluated with respect to syntactic correctness, diversity, and practical applicability, supported by the introduction of a novel metric for assessing attack diversity. Additionally, the generated attacks are validated in a controlled test environment to assess their effectiveness in realistic scenarios.
The results demonstrate that while individual generative approaches exhibit distinct strengths, their integration enables the most balanced and effective generation of synthetic cyber-attack data. The thesis concludes that combining Generative Adversarial Networks and Large Language Models provides a robust framework for producing realistic, diverse, and practically useful datasets, thereby contributing to the development of more reliable cybersecurity systems. |
|---|
| Ključne besede: | large language models, attack data generation, artificial intelligence, cybersecurity |
|---|
| Kraj izida: | Maribor |
|---|
| Kraj izvedbe: | Maribor |
|---|
| Založnik: | [B. Zgaga] |
|---|
| Leto izida: | 2026 |
|---|
| Št. strani: | 1 spletni vir (1 datoteka PDF ([XXI], 67 str.)) |
|---|
| PID: | 20.500.12556/DKUM-97077  |
|---|
| UDK: | 004.6:004.8(043.2) |
|---|
| COBISS.SI-ID: | 281650179  |
|---|
| Datum objave v DKUM: | 03.06.2026 |
|---|
| Število ogledov: | 312 |
|---|
| Število prenosov: | 13 |
|---|
| Metapodatki: |  |
|---|
| Področja: | KTFMB - FERI
|
|---|
|
:
|
Kopiraj citat |
|---|
| | | | Skupna ocena: | (0 glasov) |
|---|
| Vaša ocena: | Ocenjevanje je dovoljeno samo prijavljenim uporabnikom. |
|---|
| Objavi na: |  |
|---|
Postavite miškin kazalec na naslov za izpis povzetka. Klik na naslov izpiše
podrobnosti ali sproži prenos. |