遇见数据集
数据链接:
官方服务:

资源简介:

The CATH database is a free, publicly available online resource that provides information on the evolutionary relationships of protein domains. It provides a hierarchical domain classification of protein structures in the Protein Data Bank. Protein structures are classified using a combination of automated and manual procedures. There are four major levels in this hierarchy; Class (secondary structure classification, e.g. mostly alpha), Architecture (classification based on overall shape), Topology (fold family) and Homologous superfamily (protein domains which are thought to share a common ancestor).

CATH数据库(CATH Database)是一款免费且公开可用的在线资源,用于提供蛋白质结构域(Protein Domains)的进化关系相关信息。该数据库对蛋白质数据银行(Protein Data Bank)收录的蛋白质结构开展层级化结构域分类工作。蛋白质结构的分类流程结合了自动化操作与人工审核两种方式。该层级体系共包含四大核心层级:类别(Class,基于二级结构的分类标准,例如以α螺旋为主的结构)、架构(Architecture,依据整体空间构象划分的类别)、拓扑结构(Topology,即折叠家族)以及同源超家族(Homologous superfamily,被认为拥有共同进化祖先的蛋白质结构域)。

搜集汇总
数据集介绍
CATH 数据集图片
背景与挑战
背景概述
CATH是一个公开的蛋白质结构分类数据库,基于从蛋白质数据库(PDB)下载的结构,将蛋白质结构域按进化关系分层分类到超家族。它提供三维结构分析、蛋白质进化、功能和保守位点研究,并包含CATH-Gene3D资源的最新版本数据。
以上内容由遇见数据集搜集并总结生成
二维码
社区交流群
二维码
科研交流群
商业服务