claude-gemini-reasoning
收藏资源简介:
该数据集包含13,627个经过高度整理和去重的问题,这些问题需要深入的多步推理、复杂逻辑、高级数学、物理和编程能力。数据来源于多个高质量的推理数据集,并优先采用了由顶级模型(包括Claude Opus 4.6/4.5、Gemini 3/3.1 Pro和Claude 4.5 Sonnet)生成的回答。数据集经过严格的清理和去重处理,采用了一种称为“模型优先级排序”的系统来保留最高质量的输出。数据格式采用对话式`messages`格式,并集成了`<think>`标签,非常适合用于DeepSeek-R1风格的微调。每个助手响应都包含一个`<think>`块,其中包含逐步推理过程,后跟最终答案。该数据集适用于微调模型以实现R1风格的“思考”行为,以及处理复杂的STEM任务和代码推理问题。
This dataset comprises 13,627 highly curated and deduplicated questions that demand deep multi-step reasoning, complex logical analysis, advanced mathematical, physical, and programming capabilities. It is sourced from multiple high-quality reasoning datasets, and prioritizes responses generated by leading models including Claude Opus 4.6/4.5, Gemini 3/3.1 Pro, and Claude 4.5 Sonnet. The dataset has been subjected to rigorous cleaning and deduplication, utilizing a "model priority ranking" system to preserve the highest-quality outputs. The dataset adopts a conversational `messages` format and integrates `<think>` tags, making it ideal for DeepSeek-R1-style fine-tuning. Each assistant response contains a `<think>` block that outlines the step-by-step reasoning process, followed by the final answer. This dataset is well-suited for fine-tuning models to exhibit R1-style "thinking" behaviors, as well as addressing complex STEM tasks and code reasoning problems.
Claude-gemini-reasoning 数据集概述
数据集基本信息
- 许可证:MIT
- 语言:英语
- 标签:推理、数学、编程、逻辑、合成、代码
- 规模类别:10K < n < 100K
- 配置名称:default
- 数据文件:dataset_final.jsonl(训练集)
数据集内容与规模
- 包含 13,627 个经过高度策划和去重的问题。
- 问题需要深入的多步推理、复杂逻辑、高等数学、物理和编程能力。
- 数据经过精心清理和去重,源自多个高质量推理数据集。
- 优先采用顶级模型生成的回答,包括 Claude Opus 4.6/4.5、Gemini 3/3.1 Pro 和 Claude 4.5 Sonnet。
数据来源
数据集整合了以下来源的数据:
- https://huggingface.co/datasets/nohurry/Opus-4.6-Reasoning-3000x-filtered
- https://huggingface.co/datasets/crownelius/Opus-4.6-Reasoning-2100x-formatted
- https://huggingface.co/datasets/reedmayhew/gemini-3.1-pro-2048-reasoning-1100x
- https://huggingface.co/datasets/TeichAI/claude-4.5-opus-high-reasoning-250x
- https://huggingface.co/datasets/Roman1111111/gemini-3-pro-10000x-hard-high-reasoning
- https://huggingface.co/datasets/TeichAI/gemini-3-pro-preview-high-reasoning-1000x
- https://huggingface.co/datasets/TeichAI/claude-sonnet-4.5-high-reasoning-250x
数据处理与过滤
- 执行了去重处理。
- 当不同来源出现相似问题时,采用模型排名优先级系统保留最高质量的输出:
- Opus 4.6
- Claude 4.5 Opus
- Gemini 3.1 Pro
- Gemini 3 Pro
- Claude 4.5 Sonnet
数据格式
- 采用对话式
messages格式,并集成了<think>标签,非常适合 DeepSeek-R1 风格的微调。 - 每个条目包含用户问题(
role: "user")和助手回答(role: "assistant")。 - 助手回答包含一个
<think>块,其中是逐步推理过程,随后是最终答案。
主要用途
- 微调模型以实现 R1 风格的“思考”行为。
- 复杂 STEM 任务:高等物理和研究生水平的数学问题。
- 代码推理:逻辑密集的软件工程问题。
注意事项
- 大约 22% 的数据其推理过程被截断。建议暂时从这些示例中移除推理部分,以获取一些非推理示例。



