niamhtracey1/Synthetic-Medical-Speech-Dataset
收藏资源简介:
Synthetic Medical Speech Dataset 是一个合成音频-文本对数据集,专为开发和评估医疗领域的自动语音识别(ASR)模型而设计。该语料库包含数千个由文本到语音(TTS)系统从医学相关文本生成的短音频片段,每个片段都与其对应的转录文本配对。由于所有内容均为合成生成,该数据集不包含任何真实患者数据或个人可识别信息。数据集类型为合成音频加文本转录,领域涵盖医疗/健康护理语音识别,语言为英语(包含医学术语和对话),内容基于精选医学文本的TTS生成语音,文件格式为wav音频文件及匹配的转录文本。
Synthetic Medical Speech Dataset is a synthetic dataset of audio–text pairs designed for developing and evaluating automatic speech recognition (ASR) models in the medical domain. The corpus contains thousands of short audio clips generated from medically relevant text using a text-to-speech (TTS) system. Each clip is paired with its corresponding transcript. Because all content is synthetically produced, the dataset does not contain any real patient data or personally identifiable information. Type: Synthetic Audio + Text Transcripts, Domain: Medical / Healthcare Speech Recognition, Language: English (medical terms and dialogue), Content: TTS-generated speech from curated medical text, File Format: tuple of wavs and matching transcripts.



