ChavyvAkvar/aya_collection_language_split-central_khmer-Converted
收藏数据链接:
官方服务:
资源简介:
这个数据集包含了对话信息,每个对话包括内容和角色两个部分,同时还记录了每个对话的token数量和一些JSON格式的元数据信息。数据集被拆分为训练集,共有约357万示例。
This dataset includes conversation information, with each conversation consisting of content and role parts, as well as the number of tokens for each conversation and some metadata information in JSON format. The dataset is split into a training set with a total of about 3.57 million examples.
提供机构:
ChavyvAkvar


