Dataset con imágenes de rayos-X patológicas y de control de neumonía provocada por COVID-19 tomadas en distintos hospitales y equipos de adquisición de imagen. Todos los pacientes tenían PCR positiva
MedVAL-Bench is the first large-scale physician-validated benchmark for medical text validation, spanning 6 diverse medical tasks and containing 840 language model-generated outputs annotated by 12 ph
The objective of this study was to assess the performance of ChatGPT (GPT-4) on all items, including those with diagrams, in the Japanese National License Examination for Pharmacists (JNLEP) and compa
This dataset contains results of evaluation of performance of the new Bing (Microsoft Corporation, Redmond, Washington, USA) and Bard (Google LLC, Mountain View, California, USA) large language model-