nc33/libri_duplicate_v4
收藏官方服务:
资源简介:
该数据集是一个包含音频和文本信息的集合,其中音频和文本都有标准化和原始两种形式。每个样本都有说话者ID、文件路径、章节ID和唯一ID。数据集分为训练集,共有2000个样本,总大小约为1.17GB。
This dataset is a collection containing audio and text information, with both normalized and original forms of audio and text. Each sample has a speaker ID, file path, chapter ID, and unique ID. The dataset is split into a training set with a total of 2000 samples, with a total size of approximately 1.17GB.
提供机构:
nc33


