遇见数据集

xtz999/jordanian_arabic_speech_nasir

收藏
Hugging Face2026-05-13 更新2026-05-31 收录
官方服务:

资源简介:

该数据集是一个包含音频和文本对的数据集,音频采样率为22050Hz,主要用于音频与文本的关联任务,如语音识别或语音合成。数据集包含127070个训练样本,总大小约为24.35GB,下载大小约为23.48GB。数据以音频文件和对应文本的形式组织,适用于机器学习模型的训练。

This dataset comprises audio-text paired samples, with an audio sampling rate of 22050 Hz. It is primarily designed for audio-text association tasks, such as automatic speech recognition (ASR) or text-to-speech (TTS) synthesis. The dataset includes 127,070 training samples, with a total size of approximately 24.35 GB and a download size of roughly 23.48 GB. The data is structured as audio files paired with their corresponding text transcripts, making it well-suited for training machine learning models.

提供机构:
xtz999
二维码
社区交流群
二维码
科研交流群
商业服务