stevenmaschan/preprocessed_gigaspeech_s_subset
收藏官方服务:
资源简介:
该数据集包含文本和其他相关特征,适用于文本分类或标注任务。数据集由训练集组成,包含超过23万个样本,每个样本包括段标识符、文本内容、输入特征、标签和自定义标签。
The dataset includes text and other related features, suitable for text classification or annotation tasks. The dataset consists of a training set with over 230,000 samples, each including a segment identifier, text content, input features, labels, and custom labels.
提供机构:
stevenmaschan


