ravi242006/snm1
收藏资源简介:
--- license: cc-by-nc-sa-4.0 configs: - config_name: v0 data_files: - split: train path: v0/train-* - config_name: v1 data_files: - split: train path: v1/train-* dataset_info: - config_name: v0 features: - name: id dtype: string - name: reannotated_assistant_content dtype: string - name: problem dtype: string - name: source dtype: string - name: solution dtype: string - name: verified dtype: 'null' - name: quality_metrics dtype: 'null' splits: - name: train num_bytes: 1279431141 num_examples: 171647 download_size: 554111459 dataset_size: 1279431141 - config_name: v1 features: - name: id dtype: string - name: reannotated_assistant_content dtype: string - name: source dtype: string - name: reannotated_messages list: - name: content dtype: string - name: role dtype: string - name: messages list: - name: content dtype: string - name: role dtype: string - name: source_dataset dtype: string - name: verified dtype: 'null' - name: quality_metrics dtype: 'null' splits: - name: train num_bytes: 25783989151 num_examples: 1679162 download_size: 11128580062 dataset_size: 25783989151 --- # 🔉 𝗦𝗟𝗔𝗠 𝗹𝗮𝗯 - 𝗥𝟭-𝗗𝗶𝘀𝘁𝗶𝗹𝗹-𝗦𝗙𝗧 Dataset Lewis Tunstall, Ed Beeching, Loubna Ben Allal, Clem Delangue 🤗 and others at Hugging Face announced today that they are - 𝗼𝗽𝗲𝗻𝗹𝘆 𝗿𝗲𝗽𝗿𝗼𝗱𝘂𝗰𝗶𝗻𝗴 𝗥𝟭 🔥 We at 𝗦𝗟𝗔𝗠 𝗹𝗮𝗯 (ServiceNow Language Models) have been cooking up something as well. Inspired by Open-r1, we have decided to open source the data **stage-by-stage** to support the open source community. 𝗕𝗼𝗼𝗸𝗺𝗮𝗿𝗸 this page! **KEY DETAILS**: - ⚗️ Distilled with DeepSeek-R1-32b - 📕 Generated using Numina-math and Tulu - 🌡️ Sampled one response per prompt # 𝗦𝗖𝗛𝗘𝗗𝗨𝗟𝗘: - 🆕 [27 Jan] Release seed set of 170,000 samples - 🛑 [28 Jan] Release the unfiltered / unverified dataset ~ 2 million samples - 🟢 [TBD] Filtered and verified version to follow shortly after - 🏁 [TBD] SFT Models released **If you use our dataset, please cite us!** ``` @misc{slam-distillation-from-r1, author = {Sathwik Tejaswi Madhusudhan and Shruthan Radhakrishna and Jash Mehta and Toby Liang}, title = {Millions scale dataset distilled from R1-32b}, howpublished = {https://huggingface.co/datasets/ServiceNow-AI/R1-Distill-SFT}, publisher = {SLAM - ServiceNow Language Models Lab} year = {2025} } ```
许可证:CC BY-NC-SA 4.0 配置项: - 配置名称:v0 数据文件: - 划分方式:训练集 - 路径:v0/train-* - 配置名称:v1 数据文件: - 划分方式:训练集 - 路径:v1/train-* 数据集信息: - 配置名称:v0 特征项: - 名称:id,数据类型:字符串 - 名称:reannotated_assistant_content,数据类型:字符串 - 名称:problem,数据类型:字符串 - 名称:source,数据类型:字符串 - 名称:solution,数据类型:字符串 - 名称:verified,数据类型:空值 - 名称:quality_metrics,数据类型:空值 数据集划分: - 名称:训练集,字节数:1279431141,样本数量:171647 下载大小:554111459 数据集总大小:1279431141 - 配置名称:v1 特征项: - 名称:id,数据类型:字符串 - 名称:reannotated_assistant_content,数据类型:字符串 - 名称:source,数据类型:字符串 - 名称:reannotated_messages,数据类型:列表,列表项: - 名称:content,数据类型:字符串 - 名称:role,数据类型:字符串 - 名称:messages,数据类型:列表,列表项: - 名称:content,数据类型:字符串 - 名称:role,数据类型:字符串 - 名称:source_dataset,数据类型:字符串 - 名称:verified,数据类型:空值 - 名称:quality_metrics,数据类型:空值 数据集划分: - 名称:训练集,字节数:25783989151,样本数量:1679162 下载大小:11128580062 数据集总大小:25783989151 # 🔉 SLAM实验室 - R1-Distill-SFT数据集 Lewis Tunstall、Ed Beeching、Loubna Ben Allal、Clem Delanghe 🤗 以及Hugging Face的其他团队今日宣布,将开源复现R1模型🔥。 我们SLAM实验室(ServiceNow语言模型团队)也潜心研发了相关成果。 受Open-R1项目启发,我们决定分阶段开源本数据集,以赋能开源社区发展。 敬请收藏本页面! ## 核心细节: - ⚗️ 采用DeepSeek-R1-32b进行模型蒸馏 - 📕 基于Numina-math与Tulu框架生成数据 - 🌡️ 每个提示词仅采样一条回复 ## 发布计划: - 🆕 【1月27日】发布17万样本的种子数据集 - 🛑 【1月28日】发布约200万样本的未过滤/未验证数据集 - 🟢 【待定】后续将发布经过过滤与验证的数据集版本 - 🏁 【待定】将发布监督微调(Supervised Fine-Tuning,SFT)模型 --- 若您使用本数据集,请引用如下文献: @misc{slam-distillation-from-r1, author = {Sathwik Tejaswi Madhusudhan、Shruthan Radhakrishna、Jash Mehta、Toby Liang}, title = {基于R1-32b蒸馏的百万级规模数据集}, howpublished = {https://huggingface.co/datasets/ServiceNow-AI/R1-Distill-SFT}, publisher = {SLAM - ServiceNow语言模型实验室} year = {2025} }



