PETCI
收藏资源简介:
PETCI是一个由芝加哥大学创建的中文成语平行英语翻译数据集,包含4310条数据。该数据集通过结合人工和机器翻译的结果,旨在提高机器翻译系统及语言学习者对中文成语翻译的能力。数据集的构建过程包括从成语词典中收集翻译,以及使用Google和DeepL的机器翻译结果。PETCI的应用领域主要集中在提升机器翻译的准确性和辅助语言学习者理解成语的深层含义。
PETCI is a parallel Chinese idiom-English translation dataset developed by the University of Chicago, which consists of 4310 samples. This dataset integrates human and machine translation outputs, with the goal of improving both the performance of machine translation systems and the proficiency of language learners in translating Chinese idioms. The construction process of PETCI includes collecting translation pairs from idiom dictionaries, as well as leveraging machine translation results from Google and DeepL. The main application areas of PETCI focus on advancing the accuracy of machine translation and assisting language learners in comprehending the deep connotations of Chinese idioms.

- 1PETCI: A Parallel English Translation Dataset of Chinese Idioms芝加哥大学 · 2022年



