Bakta database
收藏资源简介:
This data repository contains the mandatory DB for Bakta. It is available in 2 versions: the default (db.tar.gz or) and a lightweight alternative (db-light.tar.gz). Bakta is a tool for the rapid & standardized local annotation of bacterial genomes & plasmids. It provides <strong>dbxref</strong>-rich and <strong>sORF</strong>-including annotations in machine-readble <code>JSON</code> & bioinformatics standard file formats for automatic downstream analysis: https://github.com/oschwengers/bakta This db provides protein sequence hash digests and lengths of UniProt's UniRef100 clusters, UniParc and NCBI RefSeq sequences for ultra-fast identification & lookups. It has been pre-annotated with several specialized db and enriched with Dbxrefs. Furthermore, seed sequences of UniProt's UniRef90 clusters are stored for fallback homology searches via Diamond sequence alignments. All conducted pre-annotations are logged and provided in the db.log.gz file. External DB versions: NCBI AMRFinderPlus: 2022-12-19.1 COG: 2020 DoriC: 12 ISFinder: 2019-09-25 Mob-suite: 2.0 Pfam: 35 RefSeq: r216 Rfam: 14.9 UniProtKB/Swiss-Prot: 2022_05 VFDB: 2023-02-10<br>



