遇见数据集

HIVMedQA

收藏
Zenodo2026-04-28 更新2026-05-26 收录
官方服务:

资源简介:

This dataset supports the findings presented in the article HIVMedQA: Benchmarking large language models for HIV medical decision support. It comprises two components: questions.csv : Contains all the questions, the corresponding gold-standard answers, and their sources. all_questions_answers_scores.csv : Includes the responses generated by LLMs, along with evaluation scores. all_questions_answers_scores_unsupervised.csv: Includes the evaluation scores under the unsupervised setting. all_questions_answers_scores_withRAG_withInfoInGuidelines.csv: Includes the responses generated by LLMs with using RAG and their associated evaluation scores. If you use this dataset in your work, please cite: Cardenal-Antolin, Gonzalo, et al. "HIVMedQA: Benchmarking large language models for HIV medical decision support." arXiv preprint arXiv:2507.18143 (2025).

提供机构:
Zenodo
创建时间:
2025-07-12
二维码
社区交流群
二维码
科研交流群
商业服务