遇见数据集

vamshirvk/numina-deepseek-r1-qwen-7b-data

收藏
Hugging Face2025-02-10 更新2025-02-15 收录
官方服务:

资源简介:

该数据集是使用distilabel工具创建的合成数据集。它包含一个pipeline.yaml文件,可用于重现生成数据集的流程。数据集包含具有问题、解决方案、消息、生成、distilabel元数据和模型名称等特征的示例。数据集被分为一个包含40个示例的训练集,数据集的总大小为813894字节。可以使用Python中的datasets库加载数据集。

This dataset is a synthetic dataset created using the distilabel tool. It includes a pipeline.yaml file that can be used to reproduce the pipeline that generated it. The dataset contains examples structured with features such as problem, solution, messages, generation, distilabel_metadata, and model_name. The dataset is split into a training set with 40 examples, and the total size of the dataset is 813894 bytes. The dataset can be loaded using the datasets library in Python.

提供机构:
vamshirvk
二维码
社区交流群
二维码
科研交流群
商业服务