Signor_3class_clustered-50
收藏资源简介:
该数据集包含结构化序列数据,主要特征包括IdA(字符串)、IdB(字符串)、SeqA(字符串)、SeqB(字符串)和labels(整型)。数据集划分为训练集(12,491个样本)、测试集(1,018个样本)和验证集(1,061个样本),总大小约20.7MB。数据文件按默认配置存储在train-*、test-*和valid-*路径下。未提供具体任务描述,但标签字段暗示可能用于分类或匹配任务。
This dataset contains structured sequential data, with core features including IdA (string), IdB (string), SeqA (string), SeqB (string), and labels (integer). The dataset is split into a training set (12,491 samples), a test set (1,018 samples), and a validation set (1,061 samples), with a total size of approximately 20.7 MB. The data files are stored under paths prefixed with train-, test-, and valid- in accordance with the default configuration. No specific task description is provided, but the labels field implies that the dataset may be used for classification or matching tasks.




