遇见数据集

fvdel05/hf_fr_dioula_full

收藏
Hugging Face2026-04-25 更新2026-05-03 收录
官方服务:

资源简介:

这是一个多语言翻译数据集,包含源语言和目标语言的文本对。数据特征包括source_lang(源语言代码)、target_lang(目标语言代码)、source(源语言文本)和target(目标语言文本)。数据集分为训练集(20513个示例)、验证集(2564个示例)和测试集(2565个示例),总大小约8.2MB,适用于机器翻译任务。

This is a multilingual translation dataset containing text pairs in source and target languages. The features include source_lang (source language code), target_lang (target language code), source (source language text), and target (target language text). The dataset is split into train (20,513 examples), validation (2,564 examples), and test (2,565 examples), with a total size of approximately 8.2 MB, suitable for machine translation tasks.

提供机构:
fvdel05
二维码
社区交流群
二维码
科研交流群
商业服务