数据链接:
官方服务:
资源简介:
custom dataset from https://www.openslr.org/109/
应用场景:
创建时间:
2023-07-14
相关数据集
MonlamAI/tibetan-voice-benchmark
这是一个藏语语音转文本的基准数据集,包含有声书、儿童语音、历史、藏语电影、自然语言、新闻、播客和藏语教学等类型的语音数据。所有的转录文本都经过了至少一个人的审核,确保了文本的准确性。数据集的转录质量分为三个等级,分别代表不同的审核程度。
Hugging Face2025-04-07 更新960
Nasiat/IUT_Regional_STT_Dataset
--- license: apache-2.0 dataset_info: features: - name: input_features sequence: sequence: float32 - name: labels sequence: int64 - name: input_length dtype: int64 splits:
Hugging Face2024-04-06 更新160
chiyuanhsiao/TTS_only_v4
该数据集包含多个特征,包括id、问题、问题语音、问题单元、指令、交错响应、文本响应和响应语音。其中,问题语音和响应语音是音频类型的数据。数据集仅包含一个训练集,共有32个样本,总大小为7657129字节,下载大小为5770828字节。
Hugging Face2024-12-11 更新90
Text2Speech_LJSpeech
# Dataset Card for "Text2Speech_LJSpeech" [More Information needed](https://github.com/huggingface/datasets/blob/main/CONTRIBUTING.md#how-to-contribute-to-the-dataset-cards)
魔搭社区2025-11-07 更新230
Nzyoka19/azure_dataset
--- license: other dataset_info: features: - name: audio dtype: audio - name: transcription dtype: string splits: - name: train num_bytes: 353754826.833 num_examples: 2573
Hugging Face2025-12-16 更新130



