HIVMedQA
收藏官方服务:
资源简介:
This dataset supports the findings presented in the article HIVMedQA: Benchmarking large language models for HIV medical decision support. It comprises two components: questions.csv : Contains all the questions, the corresponding gold-standard answers, and their sources. all_questions_answers_scores.csv : Includes the responses generated by LLMs, along with evaluation scores. If you use this dataset in your work, please cite: Cardenal-Antolin, Gonzalo, et al. "HIVMedQA: Benchmarking large language models for HIV medical decision support." arXiv preprint arXiv:2507.18143 (2025).
提供机构:
Zenodo创建时间:
2025-07-12



