遇见数据集

PYD4320/audio-text-pair_train-test-dataset-hindi

收藏
Hugging Face2024-05-30 更新2024-06-12 收录
官方服务:

资源简介:

该数据集包含音频和文本对,其中文本是印地语的转录,音频是对应的音频文件。数据集分为训练集和测试集,训练集包含200个样本,测试集包含60个样本。音频文件的采样率为16000Hz。数据集的下载大小为773258613字节,数据集总大小为783087667字节。数据集的标签包括印地语、音频、文本、音频-文本对等。

This dataset consists of audio-text pairs, where the text is the Hindi transcription and the audio is the corresponding audio file. The dataset is split into training and test sets, with 200 samples in the training set and 60 samples in the test set. All audio files have a sampling rate of 16000 Hz. The download size of the dataset is 773258613 bytes, and the total storage size is 783087667 bytes. The labels of the dataset include Hindi, audio, text, and audio-text pairs.

提供机构:
PYD4320
原始信息汇总

数据集概述

数据集特征

  • audio: 音频数据,采样率为16000 Hz。
  • text: 文本数据,类型为字符串。

数据集分割

  • train: 包含200个样本,总大小为636110732字节。
  • test: 包含60个样本,总大小为146976935字节。

数据集大小

  • 下载大小: 773258613字节。
  • 数据集总大小: 783087667字节。

数据集配置

  • default:
    • train: 数据文件路径为data/train-*
    • test: 数据文件路径为data/test-*

语言和标签

  • 语言: 印地语(Hindi)。
  • 标签:
    • hindi
    • audio
    • text
    • audio-text
    • pairs

数据集结构

  • train: 包含200行数据,特征为[audio, text]。
  • test: 包含60行数据,特征为[audio, text]。

示例数据

  • audio: 包含音频文件路径、音频数据数组及采样率。
  • text: 包含印地语文本。
二维码
社区交流群
二维码
科研交流群
商业服务