MediQAl
收藏资源简介:
MediQAl是一个法国医疗问答数据集,旨在评估语言模型在真实世界临床场景中的医学知识回忆和推理能力。该数据集包含来自法国医学考试的32,603个问题,涉及41个医学学科。数据集包括三个任务:单选题、多选题和开放式简答题。每个问题都被标记为理解或推理,以便对模型的认知能力进行详细分析。MediQAl数据集通过使用14个大型语言模型进行广泛评估,提供了一个全面的基准,用于评估语言模型在法语医学问答任务上的性能。
MediQAl is a French medical question answering dataset designed to evaluate the medical knowledge recall and reasoning capabilities of language models in real-world clinical scenarios. It contains 32,603 questions sourced from French medical examinations, covering 41 medical disciplines. The dataset includes three task types: single-choice questions, multiple-choice questions, and open-ended short answer questions. Each question is annotated as either comprehension or reasoning to enable detailed analysis of the model's cognitive abilities. The MediQAl dataset has been extensively evaluated using 14 large language models, providing a comprehensive benchmark for assessing the performance of language models on French medical question answering tasks.




