遇见数据集

EvaluateBM: a multi-agent framework for evaluating reasoning-capable language models in biomedical tasks

收藏
Zenodo2026-03-04 更新2026-05-26 收录
官方服务:

资源简介:

This is the visualization code and data of four biomedical tasks: clinical QA (MedMCQA), evidence extraction (PubMedQA), gene set functional annotation, and cell type identification for manuscript "EvaluateBM: a multi-agent framework for evaluating reasoning-capable language models in biomedical tasks"

提供机构:
Zenodo
创建时间:
2026-03-04
二维码
社区交流群
二维码
科研交流群
商业服务