NorMedQA
收藏资源简介:
NorMedQA旨在评估大型语言模型(LLMs)在挪威语境(Bokmål和Nynorsk)中的医学知识和推理能力。基准测试包含1241个问答对,涵盖多个医学领域。数据是从公开可用的挪威医学考试问题来源收集的,并经过检查、清理和预处理。
NorMedQA is designed to evaluate the medical knowledge and reasoning abilities of large language models (LLMs) within the Norwegian context (Bokmål and Nynorsk). The benchmark test consists of 1241 question-answer pairs, spanning multiple medical domains. The data is collected from publicly available Norwegian medical examination questions, and has been checked, cleaned, and preprocessed.
NorMedQA: 挪威医学问答基准与数据集概述
数据集基本信息
- 名称: NorMedQA (Norwegian Medical Question Answering Dataset)
- 语言: 挪威语(Bokmål和Nynorsk)
- 数据量: 1241个问答对
- 领域: 医学领域,涵盖多个医学专业
- 数据来源: 公开可用的挪威医学考试题目
- 数据处理: 经过检查、清理和预处理
数据集获取
- 存储位置: Zenodo
- 访问地址: https://zenodo.org/records/15345466
- 版本: 1.0
- 发布者: Riegler, M. A. (2025)
数据集特点
- 用途: 评估大型语言模型(LLMs)在挪威语境下的医学知识和推理能力
- 数据拆分: 包含将原始数据文件拆分为训练集/测试集的代码
评估指标
exact_match: 生成答案与参考答案完全匹配的百分比rouge: 基于n-grams和最长公共子序列的生成答案与参考答案重叠度测量(包括rouge1、rouge2、rougeL、rougeLsum)
使用许可
- 许可证: CC BY 4.0
引用信息
bibtex @dataset{riegler_michael_alexander_2025_15320038, author = {Riegler, Michael Alexander}, title = {{Norwegian Medical Question Answering Dataset - NorMedQA}}, month = may, year = 2025, publisher = {Zenodo}, version = {1.0}, doi = {10.5281/zenodo.15320038}, url = {https://doi.org/10.5281/zenodo.15320037} }
相关资源
- 基准测试代码库: https://github.com/kelkalot/normedqa
- Colab笔记本: https://colab.research.google.com/drive/1sDYReWYdt-3vYiAibqAohrAqTBD7aJHr?usp=sharing




