遇见数据集

MikeGreen2710/tho_cu_merged_old_new_01_10_pddvt_sanitized_old_tkn_old_ner_11000000_12000000

收藏
Hugging Face2025-10-12 更新2025-10-25 收录
官方服务:

资源简介:

该数据集包含了文本和相关标记信息,用于训练模型识别文本中的不同实体或事件。数据集分为训练集,共有100万个示例。每个示例包含一个唯一的id和文本内容,以及多个标记列表,如CIT、STR、LOC等,可能表示不同类型的实体或事件。

The dataset includes text and related annotation information for training models to identify different entities or events in the text. The dataset is split into a training set with a total of 1,000,000 examples. Each example contains a unique id, text content, and multiple annotation lists such as CIT, STR, LOC, etc., which may represent different types of entities or events.

提供机构:
MikeGreen2710
二维码
社区交流群
二维码
科研交流群
商业服务