遇见数据集

Rhyme analysis of Ukrainian Ballads: Towards a Computational Poetics

收藏
Zenodo2024-11-27 更新2026-05-26 收录
官方服务:

资源简介:

This dataset is based on the folklore collection Folk songs of Khmelnytsky region (Iefremova & Dmytrenko, 2014). The text corpus includes 200 ballads in Ukrainian language. Ballads collected in the period between 1918 and 2010. There are 9125 ballad lines in the table; 40543 tokens. To analyze the rhyming and rhythmic elements of the text corpus of Khmelnytsky region ballads, the programming language R along with RStudio was used. Code written for text analysis in Estonian Literary Museum. This dataset consist of such files: ballads_corpus_Khmelnytsky region: contains text data of Khmelnytsky region ballads in CSV; stanza+syllables: R script to analyse the stanzas types and to calculate the number of syllables in each line; finals in lines+POS analysis: R script to analyse the rhyme scheme by determine the final syllable in a line, and to do the PoS tags analysis of the rhyme; rhyme_schemes: R script to analyze the distribution of rhyme schemes in Khmelnytsky region ballads; ballads_corpus_POS: CSV file containing the text data of Khmelnytsky region ballads with part-of-speech (PoS) tags for the final word in each line; rhyme_by_PoS: R script to analyze the distribution of rhymes by part of speech across the ballads; ballads_fin_str: contains text data of Khmelnytsky region ballads with marked stress position in the last word in each line (in CSV); rhyme_stressed_position: R script to analyse rhyme by stress position.

提供机构:
Zenodo
创建时间:
2023-11-10
二维码
社区交流群
二维码
科研交流群
商业服务