遇见数据集

SayantanJoker/hindi_dataset_v2_part2_quality_metadata_description

收藏
Hugging Face2025-04-06 更新2025-04-12 收录
官方服务:

资源简介:

该数据集包含多个音频相关的特征,如文本内容、文件名、音高平均值、音高标准差、信噪比(snr)、c50、说话速率、音素、stoi、si-sdr、pesq等。数据集被划分为训练集,包含55339个示例。数据集的下载大小为6.6MB,实际大小为20.1MB。

The dataset includes various audio-related features such as text content, file name, mean utterance pitch, standard deviation of utterance pitch, signal-to-noise ratio (snr), c50, speaking rate, phonemes, stoi, si-sdr, pesq, etc. The dataset is split into a training set with 55,339 examples. The download size of the dataset is 6.6MB, and the actual size is 20.1MB.

提供机构:
SayantanJoker
二维码
社区交流群
二维码
科研交流群
商业服务