遇见数据集

Anonymised ST–MT–PE–REF Corpus for Chinese–English Institutional MTPE

收藏
Zenodo2026-06-08 更新2026-06-12 收录
官方服务:

资源简介:

This dataset contains an anonymised 420-segment ST–MT–PE–REF analytical corpus constructed for a study of machine translation post-editing in Chinese–English specialised institutional texts. The corpus includes 140 government work report segments, 140 policy document segments and 140 official news report segments. Each record includes source text, raw machine translation output, post-edited version, official reference translation, editing-operation labels, edit count, quality scores and alignment notes. The dataset is accompanied by an annotation codebook and a README file.

提供机构:
Zenodo
创建时间:
2026-06-08
二维码
社区交流群
二维码
科研交流群
商业服务