数据链接:
官方服务:
资源简介:
UMLS Heading Sequences in Spanish used to compute Word embeddings for the Spanish clinical language
应用场景:
创建时间:
2022-09-14
相关数据集
cyrille-elie/CHSA-Triage-Medic-Full-Dataset
该数据集是在AI工程师项目(CHSA项目)框架下构建的,旨在训练一个能够进行紧急分诊和提供临床推理的智能医疗助手。数据集分为三个不同的子集,对应于不同的训练阶段(监督微调和对齐)。 1. **sft_medical_dataset**:包含一般医学知识和临床案例(法语/英语),平衡为50%法语和50%英语,已通过Microsoft Presidio进行匿名化处理。 2. **sft_expert
Hugging Face2025-12-15 更新280
CLIP: A Dataset for Extracting Action Items for Physicians from Hospital Discharge Notes
We created a dataset of clinical action items annotated over MIMIC-III. This dataset, which we call CLIP, is annotated by physicians and covers 718 discharge summaries, representing 107,494 sentences.
DataCite Commons2021-12-16 更新70
Evaluation of natural language processing algorithm performance through gold-standard manual review (n = 300 clinical notes).
Evaluation of natural language processing algorithm performance through gold-standard manual review (n = 300 clinical notes).
NIAID Data Ecosystem30
SemClinBr
SemClinBr是一个多机构、多专业的葡萄牙语临床NLP任务语义标注语料库。该数据集包含1000份临床笔记,标注了65,117个实体和11,263个关系,支持多种临床NLP任务,旨在推动葡萄牙语电子健康记录的二次使用。数据集创建过程中,采用了精细的标注方案和基于网络的标注工具,确保了标注的一致性和效率。该数据集的应用领域包括临床信息提取、医学概念识别和医疗决策支持系统等。
arXiv2020-01-28 更新80
BRATECA (Brazilian Tertiary Care Dataset): a Clinical Information Dataset for the Portuguese Language
Computational medicine research requires clinical data for training and testing purposes, so the development of datasets composed of real hospital data is of utmost importance in this field. Most such
DataCite Commons2022-07-14 更新50



