台湾普通话自发语音语料库
收藏资源简介:
台湾普通话自发语音语料库是由德国蒂宾根大学定量语言学系创建的,用于研究台湾普通话中单音节词的音高轮廓。该数据集包含3824个单音节词的音高数据,涵盖63种不同的词类型。数据集的创建过程包括使用Montreal Forced Aligner进行语音对齐,并通过Praat脚本测量音高。该数据集主要用于研究音高轮廓在自发对话中的实现,以及语境和词义对音高轮廓的影响,旨在解决音高轮廓在不同语境下的变化问题。
The Spontaneous Speech Corpus of Taiwanese Mandarin was developed by the Department of Quantitative Linguistics at the University of Tübingen, Germany, to investigate the pitch contours of monosyllabic words in Taiwanese Mandarin. This corpus contains pitch data for 3824 monosyllabic words, covering 63 distinct word types. The corpus’s development process included using Montreal Forced Aligner for speech alignment and measuring pitch via Praat scripts. It is primarily used to study the realization of pitch contours in spontaneous conversations, as well as the impacts of context and lexical meaning on pitch contours, with the goal of addressing the issue of variations in pitch contours across different contexts.




