Automated Fairness Testing of Large Language Models
收藏官方服务:
资源简介:
This directory contains the evaluation data for the proposal presented in the Master's Thesis "Automated Fairness Testing of Large Language Models". Specifically, it includes the following: base-experiment/: Contains the test cases (test-cases/) generated to address RQ1 and RQ2, along with the results obtained (executions/) after running them on the models under test. stability-experiment/: Contains the test cases (test-cases/) generated to address RQ3, along with the results obtained (executions/) after running them a total of 30 times on the models under test.
提供机构:
Zenodo创建时间:
2024-10-31



