beyoru/Deepseek-v4-pro-max-distill-1000x
收藏资源简介:
该数据集包含由DeepSeek-V4-Pro(`reasoning_effort=max`,`thinking.enabled=true`)生成的推理痕迹和最终答案,使用的提示样本来自`Jackrong/GLM-5.1-Reasoning-1M-Cleaned`数据集。目标是检查质量。数据集包含1000个样本,主要语言为英语,也有一些中文/多语言STEM内容。每个样本都是一个JSON对象,包含id、domain、prompt、reasoning、response、model和usage等字段。
This dataset contains reasoning traces and final answers generated by **DeepSeek-V4-Pro** (`reasoning_effort=max`, `thinking.enabled=true`) using prompts sampled from [`Jackrong/GLM-5.1-Reasoning-1M-Cleaned`](https://huggingface.co/datasets/Jackrong/GLM-5.1-Reasoning-1M-Cleaned). Goal: just check quality. The dataset contains 1000 samples, primarily in English, with some Chinese / multilingual STEM content. Each sample is a JSON object with fields including id, domain, prompt, reasoning, response, model, and usage.




