ISSUES OF CREATING LINGUISTIC SUPPORT FOR THE UZBEKISTAN ELECTRONIC CORPS OF THE DIALECTS OF THE FERGANA REGION
收藏资源简介:
This article discusses the issues of creating a dialectal linguistic resource for the electronic corpus of the Uzbek language based on the phonetic, lexical and morphological characteristics of the dialects of the Fergana region. The study analyzes the theory of corpus linguistics, the principles of constructing a dialectal corpus, linguogeographic classification, areal localization, and the stages of entering written and audio texts into the corpus database. The principles of classifying dialectal units, forming a metadata database, and creating a search interface are substantiated using the example of the dialects of the Bogdod and Buvayda districts. The results of the study will serve to enrich the national corpus of the Uzbek language, develop research in the field of dialectology and computer linguistics, and create an important linguistic resource for speech processing systems based on artificial intelligence.



