CATH
收藏资源简介:
The CATH database is a free, publicly available online resource that provides information on the evolutionary relationships of protein domains. It provides a hierarchical domain classification of protein structures in the Protein Data Bank. Protein structures are classified using a combination of automated and manual procedures. There are four major levels in this hierarchy; Class (secondary structure classification, e.g. mostly alpha), Architecture (classification based on overall shape), Topology (fold family) and Homologous superfamily (protein domains which are thought to share a common ancestor).
CATH数据库(CATH Database)是一款免费且公开可用的在线资源,用于提供蛋白质结构域(Protein Domains)的进化关系相关信息。该数据库对蛋白质数据银行(Protein Data Bank)收录的蛋白质结构开展层级化结构域分类工作。蛋白质结构的分类流程结合了自动化操作与人工审核两种方式。该层级体系共包含四大核心层级:类别(Class,基于二级结构的分类标准,例如以α螺旋为主的结构)、架构(Architecture,依据整体空间构象划分的类别)、拓扑结构(Topology,即折叠家族)以及同源超家族(Homologous superfamily,被认为拥有共同进化祖先的蛋白质结构域)。




