遇见数据集

Labelled Hate Speech Detection Dataset.

收藏
NIAID Data Ecosystem2026-03-13 收录
官方服务:

资源简介:

This dataset contains 3000 labelled comments and posts scraped from the Reddit, Twitter and 4Chan social media websites in 2022. In this dataset, 2400 comments are labelled as non-hateful or '0', and 600 comments are labelled as hateful or '1', making an even 80/20 split. This dataset's primary purpose is for the use in machine learning classifications of hateful speech in the online speher.

创建时间:
2022-04-30
二维码
社区交流群
二维码
科研交流群
商业服务