遇见数据集

A Chinese Dataset for Evaluating the Safeguards in Large Language Models

收藏
arXiv2024-02-19 更新2024-08-06 收录
数据链接:
官方服务:

资源简介:

用于评估大型语言模型安全机制的中文数据集,旨在填补非英语语言模型安全评估的空白,并扩展到其他场景以识别风险提示拒绝的假阴性和假阳性示例。

A Chinese dataset for evaluating the safety mechanisms of large language models (LLMs), which aims to fill the gap in safety evaluation of non-English language models and is extended to other scenarios to identify false negative and false positive examples of risky prompt refusals.

创建时间:
2024-02-19
二维码
社区交流群
二维码
科研交流群
商业服务