医药大模型训练数据集
收藏资源简介:
医药大模型训练数据集包含不良反应症状及结果、药物名称、药物剂型、适应症、活性成分名称、设备分类、设备召回、设备注册、专利名称、专利申请号、专利摘要等字段(更多字段,可通过面谈沟通获取更多详细字段信息)。数据量级超过1000万条,可用于医药行业大模型训练或医药行业科研等场景。
The Pharmaceutical Large Language Model (LLM) Training Dataset encompasses fields including adverse reaction symptoms and corresponding outcomes, drug names, drug dosage forms, therapeutic indications, names of active ingredients, medical device classifications, device recalls, device registrations, patent titles, patent application numbers, and patent abstracts. Additional detailed field information is available through face-to-face consultations. The dataset comprises over 10 million entries, and is applicable for scenarios such as pharmaceutical industry LLM training and pharmaceutical scientific research.




