hausa_voa_ner
收藏官方服务:
资源简介:
Hausa VOA NER语料库专注于豪萨语的命名实体识别任务,基于VOA豪萨语新闻语料构建。该语料库包含超过1000条样本,并提供了训练集、验证集和测试集划分。数据集中,每个样本都包含文本分词和对应的命名实体标签,标签类型包括人物、组织、地点和日期等。该语料库在CC-BY 4.0许可下发布,允许用户自由使用和共享。
The Hausa VOA NER Corpus focuses on the named entity recognition (NER) task for the Hausa language, and is constructed based on the VOA Hausa news corpus. This corpus contains over 1000 samples, and provides splits for training, validation and test sets. Each sample in the dataset includes text tokens and their corresponding named entity tags, with tag types covering person, organization, location, date and other categories. This corpus is released under the CC-BY 4.0 license, allowing users to freely use and share it.
创建时间:
2024-07-19
搜集汇总
数据集介绍

背景与挑战
背景概述
Hausa VOA NER是一个用于豪萨语命名实体识别(NER)的数据集,基于VOA豪萨语新闻语料构建,包含超过1000条样本,并提供训练集、验证集和测试集的划分。该数据集标注了人物、组织、地点和日期等实体类型,以CC-BY 4.0许可发布,便于用户自由使用和共享。
以上内容由遇见数据集搜集并总结生成



