遇见数据集

医药大模型训练数据集

收藏
官方服务:

资源简介:

医药大模型训练数据集包含不良反应症状及结果、药物名称、药物剂型、适应症、活性成分名称、设备分类、设备召回、设备注册、专利名称、专利申请号、专利摘要等字段(更多字段,可通过面谈沟通获取更多详细字段信息)。数据量级超过1000万条,可用于医药行业大模型训练或医药行业科研等场景。

The Pharmaceutical Large Language Model (LLM) Training Dataset encompasses fields including adverse reaction symptoms and corresponding outcomes, drug names, drug dosage forms, therapeutic indications, names of active ingredients, medical device classifications, device recalls, device registrations, patent titles, patent application numbers, and patent abstracts. Additional detailed field information is available through face-to-face consultations. The dataset comprises over 10 million entries, and is applicable for scenarios such as pharmaceutical industry LLM training and pharmaceutical scientific research.

搜集汇总
数据集介绍
医药大模型训练数据集 数据集图片
背景与挑战
背景概述
该医药大模型训练数据集涵盖药物信息、医疗设备数据和专利文献等10余个核心字段,数据量超1000万条,适用于医疗AI模型训练及科研用途。
以上内容由遇见数据集搜集并总结生成
二维码
社区交流群
二维码
科研交流群
商业服务