相关数据集
conversations-en2bn
该数据集包含会话信息,每个会话包括ID、使用的模型、具体内容和角色、会话轮次、语言。此外,每个会话都经过OpenAI的审核,审核结果包括是否包含骚扰、威胁、仇恨、自我伤害、性内容、针对未成年人的性内容、暴力或图形暴力等分类,以及对应的分数。数据集还包含是否标记为有问题和是否编辑过的信息。数据集分为训练集,并提供了默认配置文件。
Hugging Face2025-03-24 更新390
AlanYky/offensive-no-instruction-with-symbol
--- dataset_info: features: - name: inputs dtype: string - name: target dtype: string splits: - name: train num_bytes: 2719611 num_examples: 2000 download_size: 1463462 d
Hugging Face2024-03-30 更新140
Online Abusive Attacks (OAA)
Online Abusive Attacks (OAA)数据集用于研究在对话中检测辱骂性语言。该数据集包含了对话交换(父-回复推文对),其中回复被标记为辱骂性或非辱骂性。该数据集的创建过程涉及从社交媒体平台收集对话交换,并使用文本分析和机器学习技术对回复进行标记。数据集的应用领域是社交网络内容审核和在线辱骂性语言的检测,旨在帮助社交媒体平台识别和减少有害内容,保护用户免受在线辱骂。
arXiv2025-08-18 更新220
Platform Gaslighting: A User-Centric Insight into Manipulated Realities in Online Content Moderation - interview transcripts
This is the repository concerning the interview connected to the above study, 12 interviews carried out with censored, marginalised content creators in the UK, Ireland, Italy, Australia and the USA.Th
figshare.northumbria.ac.uk2024-11-19 更新130



