遇见数据集

REVIEWER PROCESSES AND PARALLEL DICTIONARY CREATION IN ANNOTATING UZBEK DIALECTS FOR TRAINING ARTIFICIAL INTELLIGENCE MODELS ON THE EXAMPLE OF THE KHOREZM DIALECT

收藏
Zenodo2026-06-03 更新2026-06-12 收录
官方服务:

资源简介:

This study presents the methodology for annotating the Khorezm dialect, a low-resource variety of the Uzbek language, and developing parallel dictionaries for training artificial intelligence (AI) models. The Khorezm dialect differs significantly from Standard Uzbek in its phonetic, lexical, and morphological features and, as a representative of the Oghuz dialect group, exhibits considerable linguistic affinity with the Turkish language. The article examines the role of reviewer processes in improving dataset quality and analyzes the impact of parallel dictionaries on the performance of Neural Machine Translation (NMT) systems. The findings highlight the importance of systematic review procedures and high-quality parallel lexical resources for enhancing the effectiveness of AI-based language technologies for dialectal Uzbek.

提供机构:
Zenodo
创建时间:
2026-06-03
二维码
社区交流群
二维码
科研交流群
商业服务