遇见数据集

cemig-ceia/tulu-3-olmo2-sft-mixture

收藏
Hugging Face2025-02-24 更新2025-04-12 收录
官方服务:

资源简介:

这个数据集包含了对话信息,每个对话信息包括内容(content)和角色(role)两个部分。数据集被划分为训练集(train),共有938,659个示例,整个数据集的大小为1,931,946,831字节。提供了默认配置,其中指定了训练集的数据文件路径。

The dataset consists of conversation messages, each including two parts: content and role. The dataset is split into a training set (train) with a total of 938,659 examples, and the entire dataset size is 1,931,946,831 bytes. A default configuration is provided, specifying the data file path for the training set.

提供机构:
cemig-ceia
二维码
社区交流群
二维码
科研交流群
商业服务