遇见数据集

Lots-of-LoRAs/task1259_ted_translation_pl_ar

收藏
Hugging Face2025-01-01 更新2025-04-12 收录
官方服务:

资源简介:

task1259_ted_translation_pl_ar数据集是一个文本生成任务的数据集,包含了从TED演讲中提取的阿拉伯语到波兰语的翻译对。该数据集由众包方式创建,采用Apache-2.0许可。数据集分为训练集、验证集和测试集,分别包含5143、643和643个样本。每个样本包括输入文本(源语言文本)、输出文本(目标语言文本)和唯一标识符。

The task1259_ted_translation_pl_ar dataset is a text generation task dataset containing translation pairs from TED talks from Polish to Arabic. The dataset was created through crowdsourcing and is licensed under Apache-2.0. It is split into training, validation, and test sets with 5143, 643, and 643 examples respectively. Each sample includes an input text (source language text), an output text (target language text), and a unique identifier.

提供机构:
Lots-of-LoRAs
二维码
社区交流群
二维码
科研交流群
商业服务