dianavdavidson/Vaani-Hindi-majority-lg-English-with-transcript0
收藏资源简介:
该数据集是一个包含音频和相关元数据的集合,主要用于语音处理和分析。数据集特征包括音频文件、语言类型、音频时长、说话人ID、说话人已知语言、性别信息、州、地区、邮政编码、居住年限、是否有转录文本可用、转录文本内容、参考图像、说话人图像哈希和话语序列ID。数据分为训练集,包含952个样本,总大小约175.5MB,下载大小约174MB。数据集可能用于多语言语音识别、说话人识别或社会语言学分析等任务。
This dataset is a collection of audio files and associated metadata, primarily intended for speech processing and analysis. Its features include audio files, language type, audio duration, speaker ID, speaker's known language, gender information, state, region, postal code, length of residence, availability of transcribed text, transcript content, reference images, speaker image hash, and utterance sequence ID. The dataset is split into a training set containing 952 samples, with a total size of approximately 175.5 MB and a download size of around 174 MB. This dataset can be applied to tasks such as multilingual speech recognition, speaker recognition, or sociolinguistic analysis.




