AI4Protein/GO_CC
收藏官方服务:
资源简介:
GO-CC数据集是基因本体(GO)项目的细胞组分部分,包含320个标签,适用于多标签分类问题。数据集中的每个样本包含蛋白质的氨基酸序列。
The GO-CC Dataset is a part of the Gene Ontology (GO) project focusing on cellular components, containing 320 labels suitable for multi-label classification tasks. Each sample in the dataset includes the amino acid sequence of a protein.
提供机构:
AI4Protein


