TIDIGITS
收藏资源简介:
<p>This corpus contains speech which was originally designed and collected at Texas Instruments, Inc. (TI) for the purpose of designing and evaluating algorithms for speaker-independent recognition of connected digit sequences. There are 326 speakers (111 men, 114 women, 50 boys and 51 girls) each pronouncing 77 digit sequences. Each speaker group is partitioned into test and training subsets.</p><br> <p>The corpus was collected at TI in 1982 in a quiet acoustic enclosure using an Electro-Voice RE-16 Dynamic Cardiod microphone, digitized at 20kHz. The waveform files are in the NIST SPHERE format.</p><br> <p> </p><br> <p><strong>Updates</strong></p><br> <p>As of April, 2015, TIDIGITS is also available in flac compressed wav. This package is available to licensees as an additional download. Not included in this version are the folders relating to handling the shortened sphere files of the original corpus.</p></br> Portions © 1993 Trustees of the University of Pennsylvania
本语料库收录的语音数据最初由德州仪器公司(Texas Instruments, Inc.,简称TI)设计并采集,旨在研发与评估面向连接数字序列的说话人无关识别算法。该语料库共包含326名发音人,其中男性111名、女性114名、男童50名、女童51名,每位发音人录制77条连接数字序列。所有发音人按组别被划分为测试子集与训练子集。 本语料库于1982年在德州仪器的安静声学环境中采集,采用Electro-Voice RE-16动圈心形麦克风录制,采样率为20kHz。波形文件采用NIST SPHERE格式存储。 **更新说明** 截至2015年4月,TIDIGITS语料库还提供FLAC压缩格式的WAV文件包,该附加下载包仅对授权许可用户开放。本版本未包含原语料库中用于处理精简版SPHERE文件的相关文件夹。 部分内容 © 1993 宾夕法尼亚大学校董会




