Unraveling the influence of essential climatic factors on the number of tones through an extensive database of languages in China
收藏资源简介:
Code 01TextGrid.praatSegment and label the sound files in all folders under the directory 02Extract voice quality data.praatExtract voice quality parameters, including jitter, shimmer, HNR, CPP, H1-H2, H1-A1, H1-A2, and H1-A3 03Extract pitch data.praatExtract pitch data, including maximum, minimum, range, mean, upper quartile, lower quartile, pitch inter-quartile range, and median absolute deviation 04Correlation Analysis and Mantel Test.RCorrelation tests between different variables and create correlation plots. 05GAMM_Voice quality~Climate factors.RExamine the relationship between climate factors and voice quality 06GAMM_Tone ~ Voice quality.RExamine the relationship between voice quality and the number of tones 07GAMM_Tone~Climate factors.RExamine the relationship between climate factors and the number of tones 08GAMM_Pitch~Climate factors.RExamine the relationship between pitch variation, the number of tones, and climate factors Data All extracted data files are in the data folder. 1525dataset.csvThe file includes data for 1,525 language varieties with the following information: geographic location names (column A), linguistic classification and ASJP name information (columns B-E), longitude and latitude and information (columns F-G), number of tones (column H), Pitch information (columns I-J), voice quality information (columns K-R), climate information (columns S-X) Geographical distance.csvThe geographic distances between 1,525 language varieties were calculated using the Delaunay-Dijkstra method Language distance.csvThe language distances between 1,525 language varieties were calculated using the ASJP method. Specifichumiditydif.csvSpecific humidity difference dataset for the locations of 1,525 language varieties Tonedif.csvTone difference dataset among 1,525 language varieties Voice quality data extracted using different methods.csvVoice quality data for 1,115 dialectal variants, analyzed at both the lexical level and the vowel "a" level. Columns B–I present voice quality parameters extracted from the vowel, while columns J–Q provide data extracted from the lexical items.



