遇见数据集

amaniopia/flores-merged

收藏
Hugging Face2025-10-11 更新2025-10-25 收录
官方服务:

资源简介:

这是一个包含源文本(source)、目标文本(target)、源语言(src_lang)和目标语言(tgt_lang)四个字段的数据集,用于训练机器翻译模型。数据集包含一个训练集,共有29910个样本。

This dataset includes four fields: source text (source), target text (target), source language (src_lang), and target language (tgt_lang), all of which are string types. The dataset is split into a training set with a total of 29,910 samples.

提供机构:
amaniopia
二维码
社区交流群
二维码
科研交流群
商业服务