SC-Para v1.0
收藏官方服务:
资源简介:
SC-Para is a high-quality parallel corpus and evaluation benchmark for Spanish-Chinese (SC) cross-lingual natural language processing tasks, focusing on machine translation (MT) and terminology alignment. The dataset covers multiple domains including news, education, medical outreach, and government services. It includes aligned sentence pairs, terminology extractions, and benchmark tasks with train/dev/test splits, baseline scripts, and a leaderboard for community submissions. This resource addresses the scarcity of high-fidelity SC parallel data, enabling advancements in applied computational linguistics for Spanish-speaking regions (e.g., Latin America) and Chinese contexts.
创建时间:
2023-01-15



