TIDIGITS
收藏资源简介:
This corpus contains speech which was originally designed and collected at Texas Instruments, Inc. (TI) for the purpose of designing and evaluating algorithms for speaker-independent recognition of connected digit sequences. There are 326 speakers (111 men, 114 women, 50 boys and 51 girls) each pronouncing 77 digit sequences. Each speaker group is partitioned into test and training subsets. The corpus was collected at TI in 1982 in a quiet acoustic enclosure using an Electro-Voice RE-16 Dynamic Cardiod microphone, digitized at 20kHz. The waveform files are in the NIST SPHERE format. Updates As of April, 2015, TIDIGITS is also available in flac compressed wav. This package is available to licensees as an additional download. Not included in this version are the folders relating to handling the shortened sphere files of the original corpus. Portions © 1993 Trustees of the University of Pennsylvania
本语料库包含的语音数据最初由德州仪器公司(Texas Instruments, Inc.,TI)设计并采集,用于研发与评估说话人无关的连续数字序列识别算法。该语料库共有326名发音者(111名男性、114名女性、50名男童及51名女童),每名发音者录制77组连续数字序列。所有发音者群体均被划分为测试子集与训练子集。本语料库于1982年在TI公司的声学隔声录音间内采集,使用Electro-Voice RE-16动圈心形指向麦克风录制,并以20kHz采样率完成数字化。波形文件采用NIST SPHERE格式存储。更新说明:截至2015年4月,TIDIGITS语料库还可提供FLAC压缩WAV格式版本,该额外下载包仅面向已获得许可的用户开放。本版本未包含用于处理原始语料库简化版SPHERE文件的相关文件夹。部分内容©1993 宾夕法尼亚大学托管委员会

- TIDIGITS数据集首次发表,由美国国家标准与技术研究院(NIST)发布,主要用于语音识别研究。
- TIDIGITS数据集首次应用于语音识别系统的开发和评估,成为该领域的重要基准数据集之一。
- 随着语音识别技术的进步,TIDIGITS数据集被广泛用于多种语音识别算法的测试和比较。
- TIDIGITS数据集在深度学习兴起后,继续被用于验证和改进基于神经网络的语音识别模型。
- 1The DARPA TIMIT Acoustic-Phonetic Continuous Speech CorpusTexas Instruments · 1990年
- 2A Comparative Study of Speech Recognition Algorithms on the TIDIGITS DatasetUniversity of California, Berkeley · 2018年
- 3Deep Learning Approaches for Speech Recognition on the TIDIGITS DatasetMassachusetts Institute of Technology · 2020年
- 4Performance Analysis of Hidden Markov Models on the TIDIGITS DatasetStanford University · 2017年
- 5Feature Extraction Techniques for Speech Recognition on the TIDIGITS DatasetCarnegie Mellon University · 2019年



