Nationwide Speech Project
收藏资源简介:
Introduction This corpus represents part of the work of the Nationwide Speech Project (NSP) conducted by the authors at Indiana University. The purpose of the NSP was to collect a large amount of speech produced by male and female talkers representing the primary regional varieties of American English: New England, Mid-Atlantic, North, Midland, South and West. This release contains approximately 60 hours of speech or nearly one hour of speech from each of 60 white American English speakers --including five male and five female talkers from the six dialect regions -- reading words and sentences. The corpus can be used for perceptual and acoustic experiments designed to explore the role of variation in spoken language processing. Such applications include speech science experiments and sociolinguistic or sociophonetic research. Data The speakers were recruited from the Indiana University community; they were all 18-25 years old at the time of recording, had lived exclusively in one region prior to age 18, and both parents of each speaker were also raised in the same region. Further demographic information about the speakers is provided in the file talkers.txt. The materials include 102 high predictability sentences and five repetitions of each of 10 hVd words. The high predictability sentences are 5-8 words in length and the final word in each sentence is highly predictable based on the preceding semantic context. The 10 hVd words are: heed, hid, hayed, head, had, hod, hud, hoes, hood and who'd. Participants were recorded one at a time by an experimenter in a sound attenuated booth (IAC Audiometric Testing Room, Model 402). Both the experimenter and the participant sat in the sound booth during testing. During the recording session, the participant was seated in front of a ViewSonic LCD flatscreen monitor (ViewPanel VG151) which mirrored the screen of a Macintosh Powerbook G3 laptop. The participant wore a Shure head-mounted microphone (SM10A) that was positioned approximately one inch from the left corner of the talker's mouth. The microphone output was fed to an Applied Research Technology microphone tube pre-amplifier. The output gain on the pre-amplifier was adjusted by the experimenter while the participant read the Grandfather Passage as a warm-up before recording began. The output of the microphone pre-amplifier was connected to a Roland UA-30 USB audio interface which digitized the signal and transmitted it via USB ports to the laptop where each utterance was recorded in an individual AIFF 16-bit digital sound file at a sampling rate of 44.1 kHz (converted to .wav format files for this release) The experimenter held the laptop on her lap and wore headphones connected to the Roland device so that she could hear the same audio signal that inputted into the laptop for recording. Samples hpspin vowel Portions © 2003 Indiana University Research and Technology Corporation, © 2007 Trustees of the University of Pennsylvania
数据集介绍 本语料库为本文作者于印第安纳大学开展的全国语音项目(Nationwide Speech Project, NSP)的部分研究成果。该项目旨在收集代表美式英语六大主要地域变体的发音人所录制的大量语音数据,六大地域变体分别为新英格兰、中大西洋、北部、中部、南部及西部美式英语。 本次发布的数据包含约60小时的语音总量,对应60名发音人每人贡献近1小时的语音内容。所有发音人均为母语为美式英语的白人,且六大方言区域各配备5名男性与5名女性发音人,所有发音人均需朗读单词与句子。本语料库可用于设计以探究语音变体在口语语言处理中作用为目标的感知实验与声学实验,相关应用场景包括语音科学实验、社会语言学或社会语音学研究。 数据说明 发音人均从印第安纳大学校园内招募:录音时年龄均为18至25岁,18岁前仅在单一地域居住,且每位发音人的父母均在该地域长大。更多关于发音人的人口统计学信息详见talkers.txt文件。 本次录制的语料包含102句高可预测性句子,以及10个hVd类单词,每个单词均需朗读5次。高可预测性句子长度为5至8个单词,且每句的末尾单词均可通过前文的语义语境进行高度预测。10个hVd类单词分别为:heed、hid、hayed、head、had、hod、hud、hoes、hood及who'd。 所有发音人均由实验员在隔音消声室(IAC Audiometric Testing Room, Model 402)内单次单独录制。测试期间,实验员与发音人均身处隔音室内。 录音过程中,发音人坐在ViewSonic LCD平面显示器(ViewPanel VG151)前方,该屏幕可镜像显示Macintosh Powerbook G3笔记本电脑的画面。发音人佩戴Shure头戴式麦克风(SM10A),麦克风位置距发音人嘴角左侧约1英寸。麦克风输出信号被接入Applied Research Technology品牌的麦克风前置放大器。录音开始前,发音人先朗读《祖父段落》作为热身环节,实验员在此期间调整前置放大器的输出增益。 麦克风前置放大器的输出信号接入Roland UA-30 USB音频接口,该接口可将信号数字化并通过USB端口传输至笔记本电脑,所有语音片段均以单个AIFF 16位数字音频文件的形式存储,采样率为44.1 kHz(本次发布包中已转换为.wav格式文件)。实验员将笔记本电脑放置于膝上,并佩戴连接至Roland音频接口的耳机,以监听实时输入笔记本电脑用于录制的音频信号。 hVd类元音样本 本数据集部分内容 © 2003 印第安纳大学研究与技术公司,© 2007 宾夕法尼亚大学校董会。




