MikeGreen2710/tho_cu_merged_old_new_01_10_pddvt_sanitized_old_tkn_old_ner_7000000_8000000
收藏数据链接:
官方服务:
资源简介:
该数据集包含多个文本相关的特征,如id和text,以及其他字符串列表类型的特征,可能表示与文本分类或标注相关的任务。数据集划分为训练集,包含100万条示例,文件大小约为1.3GB。具体应用场景和详细描述未在README中提供。
The dataset includes various text-related features such as id and text, along with other string list features, which might indicate a text classification or annotation task. The dataset is split into a training set containing 1,000,000 examples, with a total file size of approximately 1.3GB. The specific application scenario and detailed description are not provided in the README.
提供机构:
MikeGreen2710


