libertas24/Agent-ValueBench
收藏资源简介:
Agent-ValueBench是首个用于评估代理价值的综合基准,涵盖28个价值系统、332个系统范围的价值维度、394个可执行环境和4,335个价值冲突任务。该基准旨在评估使用工具的语言模型代理的价值导向行为。每个基准案例定义了一个价值冲突任务、一个沙盒环境、可用工具以及用于评估代理轨迹是否支持价值冲突某一方的评分项。该数据集包含结构化的JSONL表格和原始基准工件,适用于在工具使用设置下评估语言模型代理的价值导向行为。
Agent-ValueBench is the first comprehensive benchmark for evaluating agent values, spanning 28 value systems, 332 system-scoped value dimensions, 394 executable environments, and 4,335 value-conflict tasks. It is designed to evaluate value-oriented behavior in tool-using language model agents. Each benchmark case defines a value-conflict task, a sandbox environment, the available tools, and rubric items used to evaluate whether an agents trajectory supports either side of the value conflict. The dataset includes structured JSONL tables for dataset viewing and Croissant metadata generation, as well as the original raw benchmark artifacts, intended for evaluating value-oriented behavior in language-model agents under tool-use settings.





