遇见数据集

moneymitrr/sample_hi_en_data_generation_omni_voice

收藏
Hugging Face2026-04-28 更新2026-05-03 收录
官方服务:

资源简介:

该数据集包含670个训练样本,每个样本具有以下特征:id(整数)、topic(字符串)、text(字符串)、audio(音频,采样率为24000)、gender(字符串)、language(字符串)、normalised_text(字符串)、normalised_number_to_text(字符串)和full_normalised_text(字符串)。数据集仅包含一个训练集,总大小为367384822字节。

The dataset contains 670 training samples, each with the following features: id (integer), topic (string), text (string), audio (with a sampling rate of 24000), gender (string), language (string), normalised_text (string), normalised_number_to_text (string), and full_normalised_text (string). The dataset includes only a training split, with a total size of 367384822 bytes.

提供机构:
moneymitrr
二维码
社区交流群
二维码
科研交流群
商业服务