AhmedDQ/dataset-tts-180
收藏官方服务:
资源简介:
该数据集是一个包含文本输入及其对应标签的训练数据集,适用于机器学习模型训练,特别是用于处理序列数据如自然语言处理任务。数据集包含三个字段:input_ids代表输入文本的编码ID,labels代表对应的标签或目标值,attention_mask用于指示输入序列中的有效位置。数据集仅包含训练集部分,共有105770个样本。
This dataset is a training dataset containing text inputs and their corresponding labels, suitable for machine learning model training, especially for sequence data processing tasks such as natural language processing. The dataset includes three fields: input_ids represent the encoded IDs of the input text, labels represent the corresponding target values, and attention_mask indicates the valid positions in the input sequence. The dataset contains only the training set, with a total of 105770 samples.
提供机构:
AhmedDQ


