遇见数据集

A dataset for evaluating clinical research claims in large language models

收藏
DataCite Commons2025-01-17 更新2025-04-20 收录
官方服务:

资源简介:

CliniFact provides a benchmark for evaluating the accuracy of large language models (LLMs) in verifying scientific claims specific to clinical research. Researchers can utilize the dataset to develop and fine-tune models to improve natural language understanding, logical reasoning, and misinformation detection in healthcare. Additionally, the dataset facilitates comparing performance across various types of LLMs.

提供机构:
figshare
创建时间:
2024-10-08
二维码
社区交流群
二维码
科研交流群
商业服务