cemig-ceia/tulu-3-olmo2-sft-mixture
收藏官方服务:
资源简介:
这个数据集包含了对话信息,每个对话信息包括内容(content)和角色(role)两个部分。数据集被划分为训练集(train),共有938,659个示例,整个数据集的大小为1,931,946,831字节。提供了默认配置,其中指定了训练集的数据文件路径。
The dataset consists of conversation messages, each including two parts: content and role. The dataset is split into a training set (train) with a total of 938,659 examples, and the entire dataset size is 1,931,946,831 bytes. A default configuration is provided, specifying the data file path for the training set.
提供机构:
cemig-ceia


