SayantanJoker/Common_voice_hindi_denoised_male_44100hz
收藏数据链接:
官方服务:
资源简介:
该数据集是一个包含音频文件及其对应转录文本的数据集,适用于语音识别等NLP任务。它包括一个训练集,共有9357个音频转录对,数据集总大小约为3.87GB。
This dataset consists of audio files and their corresponding transcriptions, suitable for NLP tasks such as speech recognition. It includes a training set with a total of 9357 audio transcription pairs, with the dataset size being approximately 3.87GB.
提供机构:
SayantanJoker


