AuAu: A Benchmark for Auditing Authoritarian Alignment in Large Language Models
收藏官方服务:
资源简介:
Data for the paper "AuAu: A Benchmark for Auditing Authoritarian Alignment in Large Language Models". Files starting with `issuebench` contain data for the realistic approach. Files starting with `final-vignette` contain results for the behavioral approach. Other files starting with `final` contain results for the psychometric approach. `final-reverse` are results for options in reverse order. `auth` or `authsys` in the filename indicates the use of the authoritarian system prompt.
提供机构:
Zenodo创建时间:
2026-02-12



