asha_twi_tts_data
收藏资源简介:
该数据集是一个语音数据集,包含音频文件及其对应的文本句子。数据集中定义了两个核心特征:path字段指向采样率为16000Hz的音频文件;sentence字段为字符串类型,存储与音频对应的文本内容。数据集被划分为训练集和测试集两部分,其中训练集包含242个样本,测试集包含46个样本。数据集总大小约为9.61MB,下载文件大小约为7.91MB。数据文件按照默认配置,从指定路径的模式文件(如train-*和test-*)加载。该数据集适用于语音识别、语音合成或音频-文本对齐等相关任务。
This dataset is a speech dataset containing audio files and their corresponding text sentences. It defines two core features: the path field points to audio files with a sampling rate of 16000Hz, and the sentence field is of string type, storing the text content corresponding to the audio. The dataset is divided into a training set and a test set, with the training set containing 242 samples and the test set containing 46 samples. The total size of the dataset is approximately 9.61MB, and the download file size is about 7.91MB. Data files are loaded according to default configurations from pattern files (such as train-* and test-*) at specified paths. This dataset is suitable for tasks such as speech recognition, speech synthesis, or audio-text alignment.
数据集概述
- 数据集名称:asha_twi_tts_data
- 数据集地址:https://huggingface.co/datasets/pius-code/asha_twi_tts_data
基本信息
该数据集是一个用于语音合成(TTS)任务的阿寒语(Twi)音频数据集,包含音频文件及对应的文本转录。
数据特征
- 音频:采样率为 16000 Hz
- 文本:句子字符串(sentence)
数据划分
| 划分 | 样本数量 | 字节数 |
|---|---|---|
| train | 256 | 10,001,007 |
| test | 46 | 1,239,693 |
数据集规模
- 下载大小:9,537,841 字节
- 总数据集大小:11,240,700 字节
配置信息
- 配置名称:default
- 训练数据文件路径:
data/train-* - 测试数据文件路径:
data/test-*




