MikeGreen2710/tho_cu_merged_old_new_01_10_pddvt_sanitized_old_tkn_old_ner_11000000_12000000
收藏数据链接:
官方服务:
资源简介:
该数据集包含了文本和相关标记信息,用于训练模型识别文本中的不同实体或事件。数据集分为训练集,共有100万个示例。每个示例包含一个唯一的id和文本内容,以及多个标记列表,如CIT、STR、LOC等,可能表示不同类型的实体或事件。
The dataset includes text and related annotation information for training models to identify different entities or events in the text. The dataset is split into a training set with a total of 1,000,000 examples. Each example contains a unique id, text content, and multiple annotation lists such as CIT, STR, LOC, etc., which may represent different types of entities or events.
提供机构:
MikeGreen2710


