CRiskEval
收藏资源简介:
CRiskEval是一个精心设计的中文数据集,用于评估LLMs中固有的风险倾向,如资源获取和恶意协调。数据集包含14,888个问题,模拟了7种风险类型的场景,每个问题附带4个由AI生成的选项,描述不同的观点或行为倾向。选项的风险级别分为极其危险、中等危险、中立和安全,并由人工修订以确保内容与风险级别匹配。
CRiskEval is a meticulously designed Chinese dataset aimed at evaluating the inherent risk tendencies in LLMs (Large Language Models), such as resource acquisition and malicious coordination. The dataset comprises 14,888 questions that simulate scenarios across seven types of risks. Each question is accompanied by four AI-generated options, depicting different perspectives or behavioral tendencies. The risk levels of the options are categorized as extremely dangerous, moderately dangerous, neutral, and safe, and have been manually revised to ensure the content aligns with the corresponding risk levels.
Risk_eval 数据集概述
数据集目的
Risk_eval 数据集专为评估大型语言模型(LLMs)在资源获取和恶意协调等方面的风险倾向性而设计。
数据集构成
- 问题数量:包含14,888个问题。
- 风险类型:模拟关于7种不同风险类型的场景。
- 选项设置:每个问题附带4个选项,这些选项由AI生成,用以描述不同的观点或行为倾向。
- 风险级别:选项内容对应的风险级别分为极其危险、中等危险、中立和安全。
- 人工校正:所有选项均经过人工修订,确保内容与风险级别相匹配。




