遇见数据集

Dataset of Khorezm dialect words of Uzbek language (Words extracted from books)

收藏
Mendeley Data2026-04-18 收录
官方服务:

资源简介:

As part of the study, a dataset was formed, which was used by a rule-oriented algorithm to standardize dialect forms into formal equivalents. In particular, the dataset contains 1340 dialect words: 1) The words in this dataset were compiled thanks to the joint work of expert linguists who are well versed not only in the Uzbek (formal) language, but also in the dialect forms of this language. 2) The sources of words in the dataset was a book, which was written by F. Abdullaev in 1965, published by the A.S. Pushkin Institute of Language and Literature of the Academy of Sciences of the Uzbek SSR.. The dataset was formed manually, no automation processes were carried out except for cases of transliteration of Cyrillic into Latin.

创建时间:
2025-05-23
二维码
社区交流群
二维码
科研交流群
商业服务