遇见数据集

DangerousQA

收藏
arXiv2025-09-30 收录
官方服务:

资源简介:

该数据集包含了200个有毒问题,覆盖了包括种族歧视、刻板印象、性别歧视和非法内容等不同有害类别。此外,该数据集旨在评估语言模型对于有毒和有害问题的回应,规模为200个问题,任务是对语言模型的安全性进行评估。

This dataset contains 200 toxic questions covering a variety of harmful categories including racial discrimination, stereotypes, gender discrimination and illegal content. Furthermore, this dataset is designed to evaluate the responses of language models to toxic and harmful questions, with a total of 200 questions, and its primary task is to assess the safety of language models.

搜集汇总
数据集介绍
DangerousQA 数据集图片
背景与挑战
背景概述
DangerousQA数据集包含200个有毒问题,覆盖种族歧视、刻板印象、性别歧视和非法内容等多种有害类别,专门用于评估语言模型对有害问题的回应,旨在测试模型的安全性。
以上内容由遇见数据集搜集并总结生成
二维码
社区交流群
二维码
科研交流群
商业服务