Repeats Proteins contact prediction based modelling datasets
收藏资源简介:
Datasets of repeats proteins model obtained by contact prediction based modelling.The multiple sequence alignments were generate using HHblits with an E-value cutoff of 0.001 against the Uniclust30_2017_04 database.Protein contacts were calculated with PconsC4 and Deep Meta Psicov were used as input for Confold. The modelling was run using the top scoring 1.5 L contracts where L is the length of the modelled regions and the two-stage modelling.For trRosetta the angles and distances were calculated with trRosetta and folded with pyRosetta.
基于接触预测建模方法构建的重复蛋白质(repeats proteins)模型数据集。本次研究中,多序列比对(multiple sequence alignments)通过HHblits工具生成,比对时以Uniclust30_2017_04数据库为检索库,设置E值阈值为0.001。蛋白质接触残基通过PconsC4计算得到,并将Deep Meta Psicov的输出结果作为Confold的输入。建模过程采用得分最高的1.5L个接触残基(其中L为待建模区域的长度),并采用两阶段建模策略。对于trRosetta模型,其结构的角度与距离参数通过trRosetta计算得到,最终使用pyRosetta完成结构折叠。



