遇见数据集

collectivat/skad_ladino_parallel

收藏
Hugging Face2025-10-20 更新2025-10-25 收录
官方服务:

资源简介:

该数据集包含 Ladino(犹太西班牙语)与英语和土耳其语对应的平行文本语料库。原始句子由SKAD创建和翻译,由Col·lectivaT进行整理。数据集分为多个配置,每个配置包含不同数量的句子对和语言组合。提供了 Ladino-英语、Ladino-土耳其语和 Ladino-西班牙语的总句数。同时包含了使用说明和引用信息。

This dataset contains parallel text corpora for Ladino (Judeo-Spanish) paired with English and Turkish. The original sentences were created and translated by SKAD and curated by Col·lectivaT. The dataset is divided into several configurations, each containing a different number of sentence pairs and language combinations. Total sentence counts for Ladino-English, Ladino-Turkish, and Ladino-Spanish are provided. Usage instructions and citation details are also included.

提供机构:
collectivat
二维码
社区交流群
二维码
科研交流群
商业服务