遇见数据集

nc33/libri_duplicate_v4

收藏
Hugging Face2025-03-30 更新2025-04-12 收录
官方服务:

资源简介:

该数据集是一个包含音频和文本信息的集合,其中音频和文本都有标准化和原始两种形式。每个样本都有说话者ID、文件路径、章节ID和唯一ID。数据集分为训练集,共有2000个样本,总大小约为1.17GB。

This dataset is a collection containing audio and text information, with both normalized and original forms of audio and text. Each sample has a speaker ID, file path, chapter ID, and unique ID. The dataset is split into a training set with a total of 2000 samples, with a total size of approximately 1.17GB.

提供机构:
nc33
二维码
社区交流群
二维码
科研交流群
商业服务