ntt123/VietBibleVox
收藏资源简介:
--- license: cc-by-sa-4.0 task_categories: - text-to-speech language: - vi pretty_name: viet-bible-vox size_categories: - 10K<n<100K --- # VietBibleVox Dataset The VietBibleVox Dataset is based on the data extracted from [open.bible](https://open.bible/) specifically for the Vietnamese language. As the original data is provided under the `cc-by-sa-4.0` license, this derived dataset is also licensed under `cc-by-sa-4.0`. The dataset comprises 29,185 pairs of (verse, audio clip), with each verse from the Bible read in Vietnamese by a male voice. - The verses are the original texts and *may not* be directly usable for training text-to-speech models. - The clips are in MP3 format with a sample rate of 48k.
VietBibleVox 数据集
概述
VietBibleVox 数据集是从 open.bible 提取的越南语数据集。该数据集基于 cc-by-sa-4.0 许可证,因此其衍生数据集也采用相同的许可证。
数据内容
- 数据集包含 29,185 对(经文,音频片段)。
- 每段经文来自圣经,由男性声音用越南语朗读。
- 经文为原文,可能不适用于直接训练文本到语音模型。
- 音频片段为 MP3 格式,采样率为 48k。
语言和大小
- 语言:越南语
- 大小:10K<n<100K
任务类别
- 文本到语音
许可证
- cc-by-sa-4.0



