遇见数据集

aimaralab/aimaralab

收藏
Hugging Face2025-11-06 更新2025-11-15 收录
官方服务:

资源简介:

这是一个由AiMara Lab创建的Aymara-Spanish平行语料库,基于manythings.org的Spanish-English平行语料库(来源于Tatoeba)。该语料库经过母语为Aymara的人手动校对,使用Google Translate模型生成,并由人类验证者进行后期编辑。此资源的目的是促进NLP工具和针对Aymara语言的自动翻译模型的开发。

This is an Aymara-Spanish parallel corpus created by AiMara Lab, based on the manythings.org Spanish-English parallel corpus (derived from Tatoeba). The corpus has been manually reviewed by native Aymara speakers, generated using the Google Translate model, and post-edited by human validators. The purpose of this resource is to promote the development of NLP tools and automatic translation models for the Aymara language.

提供机构:
aimaralab
二维码
社区交流群
二维码
科研交流群
商业服务