遇见数据集

reddyrohith49471/jt-dataset-final2

收藏
Hugging Face2026-04-27 更新2026-05-03 收录
官方服务:

资源简介:

该数据集是一个多语言语音数据集,包含音频文件及其对应的文本句子,每个样本带有说话者ID和语言标签。音频采样率为16000Hz,数据集分为训练集(3582个样本)和测试集(398个样本),总大小约为681MB。

This dataset is a multilingual speech dataset containing audio files with corresponding text sentences, each labeled with speaker ID and language. The audio sampling rate is 16000Hz, and the dataset is split into train (3582 examples) and test (398 examples) sets, with a total size of approximately 681MB.

提供机构:
reddyrohith49471
二维码
社区交流群
二维码
科研交流群
商业服务