官方服务:
资源简介:
Word embeddings obtained from clinical data in Spanish.
应用场景:
创建时间:
2019-12-12
相关数据集
MIMIC-IV-Note: Deidentified free-text clinical notes
The advent of large, open access text databases has driven advances in state- of-the-art model performance in natural language processing (NLP). The relatively limited amount of clinical data availabl
DataCite Commons2024-12-22 更新550
Clinical-T5: Large Language Models Built Using MIMIC Clinical Text
Recent advances in scaling large language models (LLMs) has resulted in significant improvements over a number of natural language processing benchmarks. There has been some work to pretrain these lan
DataCite Commons2023-01-25 更新120
BRATECA (Brazilian Tertiary Care Dataset): a Clinical Information Dataset for the Portuguese Language
Computational medicine research requires clinical data for training and testing purposes, so the development of datasets composed of real hospital data is of utmost importance in this field. Most such
DataCite Commons2022-07-14 更新50
cyrille-elie/CHSA-Triage-Medic-Full-Dataset
该数据集是在AI工程师项目(CHSA项目)框架下构建的,旨在训练一个能够进行紧急分诊和提供临床推理的智能医疗助手。数据集分为三个不同的子集,对应于不同的训练阶段(监督微调和对齐)。 1. **sft_medical_dataset**:包含一般医学知识和临床案例(法语/英语),平衡为50%法语和50%英语,已通过Microsoft Presidio进行匿名化处理。 2. **sft_expert
Hugging Face2025-12-15 更新280
CLIP: A Dataset for Extracting Action Items for Physicians from Hospital Discharge Notes
We created a dataset of clinical action items annotated over MIMIC-III. This dataset, which we call CLIP, is annotated by physicians and covers 718 discharge summaries, representing 107,494 sentences.
DataCite Commons2021-12-16 更新70



