遇见数据集

HASSANIYA-DTCD: A new Dataset for Benchmarking Text Classification Tasks on HASSANIYA Dialect

收藏
Mendeley Data2026-04-18 收录
官方服务:

资源简介:

HASSANIYA-DTCD: A new Dataset for Benchmarking Text Classification Tasks on HASSANIYA dialect is the first Mauritanian dialect dataset called “HASSANIYA” containing 1851 records classified into three categories: positive, negative, and neutral. This dataset was collected using web scraping tools from comments posted on the Facebook platform, and Label Studio was used to annotate each record. For more details, see the README file.

创建时间:
2025-05-06
二维码
社区交流群
二维码
科研交流群
商业服务