IIC/livingner1
收藏官方服务:
资源简介:
LivingNER数据集是一个专注于生物医学和临床领域的西班牙语数据集,主要用于命名实体识别(NER)任务。该数据集仅包含西班牙语的任务1数据,不包括多语言数据和背景数据。它是论文中基准测试的一部分,并提供了原始数据集的引用信息。
The LivingNER dataset is a Spanish-language dataset focused on the biomedical and clinical domains, primarily used for named entity recognition (NER) tasks. This dataset only includes Task 1 data in Spanish, with no multilingual data or background data provided. It is part of the benchmark tests in the associated paper, and the citation information of the original dataset is provided.
提供机构:
IIC原始信息汇总
数据集概述
基本信息
- 名称: LivingNER
- 语言: 西班牙语(es)
- 标签: 生物医学、临床、西班牙语
- 多语言性: 单语
- 任务类别: 词元分类
- 任务ID: 命名实体识别
- 许可证: CC-BY-4.0
训练与评估
- 任务: 词元分类
- 训练与评估分割:
- 训练分割: train
- 评估分割: test
- 评估指标: F1分数
数据集内容
- 仅包含西班牙语的任务1数据,不包括多语言数据和背景数据。
引用信息
- 原始数据集引用: bibtex @article{amiranda2022nlp, title={Mention detection, normalization & classification of species, pathogens, humans and food in clinical documents: Overview of LivingNER shared task and resources}, author={Miranda-Escalada, Antonio and Farr{e}-Maduell, Eul{`a}lia and Lima-L{o}pez, Salvador and Estrada, Darryl and Gasc{o}, Luis and Krallinger, Martin}, journal = {Procesamiento del Lenguaje Natural}, year={2022} }



