遇见数据集

SayantanJoker/audio_hindi_karya_tts_description_8

收藏
Hugging Face2025-03-31 更新2025-04-12 收录
官方服务:

资源简介:

该数据集包含了语音相关的多个特征,如文本内容、文件名、音高平均值、音高标准差、信噪比、C50值、说话速率、音素信息、STOI值、SI-SDR值、PESQ值、噪声情况、混响情况、语音单调性、SDR噪声和PESQ语音质量等。数据集分为训练集,共有9978个样本。数据集的总大小为4019960字节,下载大小为1317496字节。

The dataset includes various speech-related features such as text content, file name, mean utterance pitch, standard deviation of pitch, signal-to-noise ratio, C50 value, speaking rate, phoneme information, STOI value, SI-SDR value, PESQ value, noise condition, reverberation, speech monotony, SDR noise, and PESQ speech quality. The dataset is split into a training set with a total of 9978 samples. The total size of the dataset is 4019960 bytes, and the download size is 1317496 bytes.

提供机构:
SayantanJoker
二维码
社区交流群
二维码
科研交流群
商业服务