Anonymised ST–MT–PE–REF Corpus for Chinese–English Institutional MTPE
收藏官方服务:
资源简介:
This dataset contains an anonymised 420-segment ST–MT–PE–REF analytical corpus constructed for a study of machine translation post-editing in Chinese–English specialised institutional texts. The corpus includes 140 government work report segments, 140 policy document segments and 140 official news report segments. Each record includes source text, raw machine translation output, post-edited version, official reference translation, editing-operation labels, edit count, quality scores and alignment notes. The dataset is accompanied by an annotation codebook and a README file.
提供机构:
Zenodo创建时间:
2026-06-08



