Evaluating statistical multiple sequence alignment in comparison to other alignment methods on protein data sets
收藏DataONE2020-06-24 更新2025-06-14 收录
下载链接:
https://search.dataone.org/view/sha256:f8141b82efcbb435c01289a63447aa97d176735273ccdaa282a3a701ed810cba
下载链接
链接失效反馈官方服务:
资源简介:
The estimation of multiple sequence alignments of protein sequences is a basic step in many bioinformatics pipelines, including protein structure prediction, protein family identification, and phylogeny estimation. Statistical co-estimation of alignments and trees under stochastic models of sequence evolution has long been considered the most rigorous technique for estimating alignments and trees, but little is known about the accuracy of such methods on biological benchmarks. We report the results of an extensive study evaluating the most popular protein alignment methods as well as the statistical co-estimation method BAli-Phy on 1192 protein data sets from established benchmarks as well as on 120 simulated data sets. Our study (which used more than 230 CPU years for the BAli-Phy analyses alone) shows that BAli-Phy has better precision and recall (with respect to the true alignments) than the other alignment methods on the simulated data sets, but has consistently lower recall on...
创建时间:
2025-06-12



