It contains the dataset used for the experimentation. Specifically, there are two text files, each containing 32000 different domain names. One file is clean domain names, the other file contains DGA
Datasets used to evaluate a proposed DGA detection approach using Integrated transformer embeddings from large language models.The original full dataset is available from:GitHub - chrmor/DGA_domains_d
This repository contains a large dataset for the research of domain generation algorithms (DGAs) and machine learning. At the time of writing the dataset contains more than 90m of domains and more