ChavyvAkvar/aya_collection_language_split-indonesian-Converted
收藏数据链接:
官方服务:
资源简介:
该数据集包含了对话信息,每个对话包括内容(content)和角色(role)信息。同时记录了每个对话的令牌数量和可能的JSON格式元数据。数据集分为训练集,共有3610078个对话示例,总大小为2.15GB。
The dataset consists of conversation information, each conversation includes content (content) and role (role) information. It also records the number of tokens in each conversation and possible JSON-formatted metadata. The dataset is split into a training set with a total of 3,610,078 conversation examples, totaling 2.15GB in size.
提供机构:
ChavyvAkvar


