遇见数据集

ferrazzipietro/LS_Mistral-7B-v0.1_ncbi_disease_NoQuant_16_64_0.01_16_BestF1

收藏
Hugging Face2024-07-08 更新2024-07-22 收录
官方服务:

资源简介:

该数据集主要用于命名实体识别任务,特别是疾病相关的实体识别。数据集包含多个特征字段,如id、tokens、ner_tags、input_ids、attention_mask、labels、sentence、predictions和ground_truth_labels。其中,ner_tags字段用于标注疾病相关的实体,包含O、B-Disease和I-Disease三个类别。数据集仅包含一个测试分割,共有924个示例,总大小为1369252字节。

This dataset is primarily used for natural language processing tasks, particularly Named Entity Recognition (NER). It includes multiple features such as id, tokens, ner_tags, input_ids, attention_mask, labels, sentence, predictions, and ground_truth_labels. Among these, tokens and ner_tags are sequence data, with ner_tags containing specific class labels like O, B-Disease, and I-Disease. The dataset is divided into a test set with 924 samples, totaling 1369252 bytes.

提供机构:
ferrazzipietro
二维码
社区交流群
二维码
科研交流群
商业服务