DIRHA-ENGLISH
收藏资源简介:
DIRHA-ENGLISH数据集是由Bruno Kessler基金会创建的多麦克风语料库,专注于家庭环境中的远场语音识别。该数据集包含真实和模拟材料,涵盖了12位美国和12位英国英语母语者的多种语音内容,如丰富的音素句子、报纸文章、对话语音、关键词和命令。数据集通过在家庭环境中分布大量麦克风和麦克风阵列来收集,旨在解决远场语音识别中的挑战,如非平稳噪声和声学混响的影响。
The DIRHA-ENGLISH dataset is a multi-microphone corpus created by the Bruno Kessler Foundation, focusing on far-field speech recognition in home environments. This dataset includes both real and simulated materials, covering diverse speech content from 12 native American English speakers and 12 native British English speakers, such as rich phonetic sentences, newspaper articles, conversational speech, keywords and commands. It was collected by deploying a large number of microphones and microphone arrays in home environments, aiming to address the core challenges in far-field speech recognition, including the impacts of non-stationary noise and acoustic reverberation.
- 1The DIRHA-English corpus and related tasks for distant-speech recognition in domestic environmentsBruno Kessler基金会 · 2017年



