Lots-of-LoRAs/task315_europarl_sv-en_language_identification
收藏数据链接:
官方服务:
资源简介:
task315_europarl_sv-en_language_identification数据集是一个用于语言识别的任务数据集,包含了瑞典语到英语的双语平行语料。数据集由众包方式创建,采用Apache-2.0许可。它包括训练集、验证集和测试集,分别包含5199、650和650个样本。数据集的特征包括输入文本、输出文本和唯一标识符。
The task315_europarl_sv-en_language_identification dataset is a language identification task dataset containing Swedish to English bilingual parallel corpora. The dataset is created through crowdsourcing and is licensed under Apache-2.0. It includes a training set, a validation set, and a test set with 5199, 650, and 650 samples respectively. The features of the dataset include input text, output text, and a unique identifier.
提供机构:
Lots-of-LoRAs


