相关数据集
MushanW/GLOBE_V2
GLOBE是一个高质量的英语语料库,包含全球各地的口音,专门设计用于解决当前零样本说话者自适应文本到语音(TTS)系统在适应有口音的说话者时表现不佳的问题。与常用的英语语料库(如LibriTTS和VCTK)相比,GLOBE独特之处在于其包含了来自23,519名说话者的语音,覆盖了全球164种口音,并提供了这些说话者的详细元数据。与原始语料库(即Common Voice)相比,GLOBE通过严格的过
Hugging Face2024-11-24 更新200
Procit004/combinedDatasetZipandHouse
该数据集包含音频和文本数据,特征字段包括说话者ID、口音、文本、性别和音频。数据集包含一个训练分割,共有300个样本,占用129594394.0字节。数据集的下载大小为118952741字节,总大小为129594394.0字节。
Hugging Face2024-07-18 更新90
Data_Sheet_1_From first encounters to longitudinal exposure: a repeated exposure-test paradigm for monitoring speech adaptation.pdf
Perceptual difficulty with an unfamiliar accent can dissipate within short time scales (e.g., within minutes), reflecting rapid adaptation effects. At the same time, long-term familiarity with an acce
NIAID Data Ecosystem20
ACLA1-FM021_B_video - SS and SE are sitting down with LD. SE has the microphone. This is a good relaxe...
SS and SE are sitting down with LD. SE has the microphone. This is a good relaxed session. SE uses quite a lot of SAE when she plays the doctor. She also uses an American accent.. Language as given:
Research Data Australia120
Data for: Cognitive processes underlying spoken word recognition during soft speech
Time series data for two Visual World Paradigm experiments. Stimuli are words presented at a conversational level (65 dBA) and words presented at lower intensities (40 dBA and 50dBA). Data files are .
NIAID Data Ecosystem50



