Poisoned SLM Ablation Dataset
收藏资源简介:
This dataset contains experimental artifacts used for conducting ablation studies on poisoned Small Language Models (SLMs). The objective is to analyze how different components of transformer architectures contribute to adversarial behaviors such as backdoor activation, data leakage, and prompt manipulation. The dataset supports the findings presented in the associated research work on LLM vulnerability assessment. The dataset is organized into four primary experimental configurations: - Model_Deep_DOS- Model_Deep_Low- Model_Shallow_High- Model_Shallow_Low The dataset includes experiments across multiple ablation strategies: - Neuron-level ablation (true / random)- Layer-wise ablation- Residual channel ablation- Mask-based ablation- Intra-block : Layer Zero sub structure ablation of Model_shallow_high



