遇见数据集

Lots-of-LoRAs/task1252_ted_translation_it_gl

收藏
Hugging Face2025-01-05 更新2025-04-12 收录
官方服务:

资源简介:

task1252_ted_translation_it_gl是一个文本生成任务的数据集,包含从TED演讲中翻译成意大利语和加里西亚语的文本对。该数据集通过众包方式收集,并分为训练集、验证集和测试集,分别包含5119、640、640个样本。数据集的目的是用于训练和评估自然语言处理模型在文本生成任务上的性能。

task1252_ted_translation_it_gl is a text generation task dataset containing text pairs translated from TED Talks into Italian and Galician. The dataset is collected through crowdsourcing and is divided into training, validation, and test sets with 5119, 640, and 640 samples respectively. The purpose of the dataset is to train and evaluate natural language processing models on their performance in text generation tasks.

提供机构:
Lots-of-LoRAs
二维码
社区交流群
二维码
科研交流群
商业服务