台湾普通话自发语音语料库
收藏资源简介:
台湾普通话自发语音语料库是由德国蒂宾根大学定量语言学系创建的,用于研究台湾普通话中单音节词的音高轮廓。该数据集包含3824个标记了音高值的单音节词,涵盖63种不同的词类型。数据集的创建过程包括使用Montreal Forced Aligner进行语音对齐,并通过Praat脚本测量音高。该数据集主要用于研究音高轮廓在自发对话中的实现,以及语境和词义对音高轮廓的影响,旨在解决音高轮廓在不同语境下的变化问题。
Taiwan Mandarin Spontaneous Speech Corpus was developed by the Department of Quantitative Linguistics, University of Tübingen, Germany, for investigating pitch contours of monosyllabic words in Taiwan Mandarin. This corpus contains 3,824 monosyllabic words annotated with pitch values, covering 63 distinct word types. Its construction process includes speech alignment using Montreal Forced Aligner and pitch measurement via Praat scripts. It is primarily used to study the realization of pitch contours in spontaneous conversations, as well as the impacts of context and word meaning on pitch contours, aiming to address the variation of pitch contours across different contexts.

- 1A corpus-based investigation of pitch contours of monosyllabic words in conversational Taiwan Mandarin德国蒂宾根大学定量语言学系 · 2024年



