OmniPath_2class_clustered-30
收藏资源简介:
该数据集包含多个特征字段,包括标识符(IdA、IdB)、标签(labels)、共识标志(consensus_stim、consensus_inhib、consesus_direction)、来源(sources)、类型(type)以及去除引用的参考文献(references_stripped)。数据集分为训练集、测试集和验证集,其中训练集包含74,542个样本,测试集包含25,476个样本,验证集包含25,000个样本。总数据集大小为11,944,548字节,下载大小为1,420,734字节。数据文件按分割存储,路径分别为data/train-*、data/test-*和data/valid-*。
This dataset includes multiple feature fields, namely identifiers (IdA, IdB), labels, consensus flags (consensus_stim, consensus_inhib, consesus_direction), sources, type, and citation-stripped references (references_stripped). The dataset is partitioned into training, test, and validation subsets: the training subset contains 74,542 samples, the test subset contains 25,476 samples, and the validation subset contains 25,000 samples. The total size of the complete dataset is 11,944,548 bytes, with a download size of 1,420,734 bytes. All data files are stored per subset, with their respective paths being data/train-*, data/test-*, and data/valid-*.



