CHAracterizing INdividual Speakers (CHAINS)
收藏资源简介:
Introduction CHAINS was created by researchers at University College Dublin and contains recordings of thirty-six English speakers reading fables and selected sentences in different speaking styles. The data was obtained in two different sessions with a time separation of about two months. The goal of the corpus is to provide a range of speaking styles and voice modifications for speakers sharing the same accent. Other existing corpora, in particular CSLU Speaker Recognition Version 1.1, TIMIT and the IViE corpus (English Intonation in the British Isles), served as referents in the selection of material. This design decision was made to ensure that methods designed and evaluated on the CHAINS corpus might be directly testable on these other corpora, which were recorded using quite different dialects and channel characteristics. Additional documentation about the corpus and its methodolgy is available at the CHAINS website. Data The data was collected in two recording sessions in a total of six different speaking styles. The first recording session was carried out in a professional recording studio in December 2005. Speakers were recorded in a sound-attenuated booth reading text in the solo, synchronous and retell styles using a Neumann U87 condenser microphone. Additional tracks using other microphones (near and far-field) were also recorded and may be made available upon request to the authors. The second recording session took place from March 2006 to May 2006 in a quiet office environment, using an AKG C420 headset condenser microphone. Speakers read text in the rsi, whisper and fast modes. The six different speaking styles were: solo reading synchronous reading spontaneous speech (retell) reptitive synchronous imitation (rsi) whispered fast reading fast speech reading In two of the speaking conditions adopted, speakers modified their speech in a constrained fashion towards a known target in the synchronous condition, the speech of the co-speaker served as a target, while in rsi, there was an explicit known static target. The presence of a known target which speakers aim to copy raises the bar in the discovery and design of procedures for automatic speaker identification, as the target speech provides a potentially highly confusing foil. The whisper and fast speech conditions are also well defined speaking styles which require substantial voice modification by the speaker. Participants were recruited through the University College Dublin and were paid for their participation. No participant had any known speech or hearing deficit. The speakers were from the United Kingdom, the eastern part of Ireland (Dublin and adjacent counties) and the United States. Further information about the speakers, their gender and dialect is available in the documentation released with this corpus. Samples For the example of the data in this particular corpus please examine this sound file of the fast reading type Portions © 2005, 2006 University College Dublin, © 2008 Trustees of the University of Pennsylvania
### 数据集简介 本数据集由都柏林大学学院(University College Dublin)的研究者构建,包含36名英语使用者以不同言语风格朗读寓言与精选语句的录音。该数据分两次采集,两次会话间隔约两个月。本语料库的构建目标是为拥有相同口音的说话者提供多样化的言语风格与语音调整方案。现有其他语料库,尤其是CSLU说话人识别版本1.1(CSLU Speaker Recognition Version 1.1)、TIMIT以及IViE语料库(不列颠群岛英语语调语料库,English Intonation in the British Isles),被用作选材时的参照标准。这一设计旨在确保基于本语料库开发与评估的方法,可直接在上述其他语料库上进行测试——尽管这些语料库采用了截然不同的方言与录音通道特性。有关本语料库及其构建方法的更多文档,可通过CHAINS官方网站获取。 ### 数据采集详情 数据采集于两次录音会话,涵盖共计6种不同的言语风格。第一次录音会话于2005年12月在专业录音棚内完成:受试者在隔音棚内使用诺依曼U87电容麦克风(Neumann U87 condenser microphone),以单人朗读、同步跟读与复述三种风格朗读文本。此外还使用其他麦克风(近场与远场)录制了额外音轨,可根据作者要求提供。第二次录音会话于2006年3月至2006年5月在安静的办公环境中完成,使用AKG C420头戴式电容麦克风(AKG C420 headset condenser microphone)。受试者以rsi(重复式同步模仿,repetitive synchronous imitation)、耳语与快速三种模式朗读文本。 六种言语风格分别为:单人朗读、同步跟读、即兴复述(retell)、重复式同步模仿(rsi)、耳语朗读、快速朗读。 在两种特定的发音场景中,受试者需按照明确目标约束调整语音:在同步跟读场景中,合作说话者的语音作为模仿目标;而在rsi场景中,则存在一个明确已知的静态语音目标。这种需要受试者模仿已知目标的设定,提升了自动说话人识别方法的研发与评估难度——因为目标语音本身会带来极具混淆性的干扰源。耳语与快速语音场景同样属于边界清晰的言语风格,要求受试者对自身语音进行大幅调整。 受试者通过都柏林大学学院招募,并会获得参与报酬。所有受试者均无已知的言语或听力障碍。参与的说话者来自英国、爱尔兰东部地区(都柏林及周边郡)与美国。有关受试者的性别、方言等更多信息,可参见本语料库附带的官方文档。 ### 数据示例 如需查看本语料库的数据示例,请参阅这份快速朗读类型的音频文件。 版权声明:部分内容 © 2005、2006 都柏林大学学院,© 2008 宾夕法尼亚大学校董会




