遇见数据集

sbhikha/Inkuba_isizulu_dev_sliced_v1

收藏
Hugging Face2024-08-24 更新2024-12-14 收录
官方服务:

资源简介:

该数据集可能用于多语言文本处理或机器翻译任务,包含指令、输入、目标输出、输入语言、目标语言以及相关的置信度分数。数据集包含一个训练集分割,共有2,019,865个样本,总大小为420,759,837字节。

This dataset is likely used for multilingual text processing or machine translation tasks, containing instructions, inputs, target outputs, input language, target language, and associated confidence scores. The dataset includes a training split with 2,019,865 samples and a total size of 420,759,837 bytes.

提供机构:
sbhikha
二维码
社区交流群
二维码
科研交流群
商业服务