遇见数据集

regmibijay/ops-volltext-klassifizierung-v2

收藏
Hugging Face2025-03-27 更新2025-04-12 收录
官方服务:

资源简介:

这是一个用于OPS分类的合成数据集,包含了所有OPS代码(包括成像代码)。数据集使用Q4_K_M量化的Qwen2.5 7B模型创建,每行的序列长度比V1版本更长。通过质量控制在LLM中减少了幻觉,从而相比V1版本有显著的质量提升。数据集的目的是支持德国医学界创建高度专业化的德语模型。

This is a synthetic dataset for OPS classification, containing all OPS codes (including imaging codes). The dataset was created using the Q4_K_M quantized Qwen2.5 7B model, with a sequence length per line longer than that of version V1. Hallucinations in LLM have been reduced through quality control, resulting in a significant quality improvement over version V1. The purpose of the dataset is to support the German medical community in creating highly specialized German models.

提供机构:
regmibijay
二维码
社区交流群
二维码
科研交流群
商业服务