遇见数据集

An error budget for political-bias measurement in large language models: data and code

收藏
Zenodo2026-09-25 更新2026-10-01 收录
官方服务:

资源简介:

Raw per-trial data, scoring adapters, and statistical analysis code for "An error budget for political-bias measurement in large language models." Nineteen large language models were administered up to four standardized political-orientation instruments (Political Compass, 8Values, SapplyValues, Pew Research Center's 2026 Political Typology) repeatedly, alongside factorial designs varying asker identity, instrument contamination, and response format, decomposing a measured political position into variance from model identity, instrument choice, and trial noise. Includes: raw per-trial administrations (data/), authoritative scoring adapters and collection scripts (scoring/), and the statistical-analysis pipeline that reproduces every number reported in the paper from this data.

提供机构:
Zenodo
创建时间:
2026-09-25
二维码
社区交流群
二维码
科研交流群
商业服务