遇见数据集

mteb/told-br

收藏
Hugging Face2025-05-04 更新2025-04-08 收录
官方服务:

资源简介:

这是一个包含文本和与之相关的六个分类标签的数据集,用于检测文本中的不同类型的偏见和不适当的内容。分类标签包括:同性恋恐惧症、粗俗、侮辱、种族主义、厌女症和排外主义,每个标签都有四种不同的投票等级。数据集包含一个训练集,共有21000个文本示例。

This dataset includes text and six associated categorical labels for detecting different types of biases and inappropriate content in the text. The categorical labels include: homophobia, obscenity, insult, racism, misogyny, and xenophobia, each with four different voting levels. The dataset contains a training set with a total of 21,000 text examples.

提供机构:
mteb
二维码
社区交流群
二维码
科研交流群
商业服务