Dataset Composition: Examples: 200,000 Benign: 100,000 Malicious: 100,000 Features: 28 Dataset composed of stratified random samples of benign domains derived from the Majestic Million list and malic
Dataset used in "Selvi, J., Rodríguez, R. J., & Soria-Olivas, E. (2019). Detection of algorithmically generated malicious domain names using masked N-grams. Expert Systems with Applications, 124, 156-
Datasets used to evaluate a proposed DGA detection approach using Integrated transformer embeddings from large language models.The original full dataset is available from:GitHub - chrmor/DGA_domains_d
The dataset is meant for supervised machine learning based analysis of malicious and non-malicious domain names. The dataset was created from scratch, using publicly DNS logs of both malicious and n