遇见数据集

sjdata/single_speaker_en_test_librivox

收藏
Hugging Face2023-07-15 更新2024-03-04 收录
官方服务:

资源简介:

--- dataset_info: features: - name: audio dtype: audio: sampling_rate: 16000 - name: text dtype: string - name: normalized_text dtype: string splits: - name: train num_bytes: 20226057306.427 num_examples: 139411 download_size: 1857190033 dataset_size: 20226057306.427 --- # Dataset Card for "single_speaker_en_test_librivox" # Created for testing, not suggested for production #### Dataset Summary The corpus consists of a single speaker extracted frrom LibriVox audiobook. #### Languages The audio is in English. #### Source Data Initial Data Collection and Normalization The voices used in my Datasets are volenteers who have donated their time and voices to open source LibriVox projects. Please respect their privacy. #### Licensing Information MIT [More Information needed](https://github.com/huggingface/datasets/blob/main/CONTRIBUTING.md#how-to-contribute-to-the-dataset-cards)

提供机构:
sjdata
原始信息汇总

数据集概述

数据集名称

single_speaker_en_test_librivox

数据集特征

  • audio:
    • 数据类型: 音频
    • 采样率: 16000 Hz
  • text:
    • 数据类型: 字符串
  • normalized_text:
    • 数据类型: 字符串

数据集分割

  • train:
    • 样本数量: 139411
    • 数据大小: 20226057306.427 字节

数据集大小

  • 下载大小: 1857190033 字节
  • 数据集总大小: 20226057306.427 字节
二维码
社区交流群
二维码
科研交流群
商业服务