SH1041816.10FU
收藏资源简介:
UNITE provides a unified way for delimiting, identifying, communicating, and working with DNA-based Species Hypotheses (SH). All fungal ITS sequences in the international nucleotide sequence databases are clustered to approximately the species level by applying a set of dynamic distance values (<0.5 - 3.0%). All species hypotheses are given a unique, stable name in the form of a DOI, and their taxonomic and ecological annotations are verified through distributed, web-based third-party annotation efforts. SHs are connected to a taxon name and its classification as far as possible (phylum, class, order, etc.) by taking into account identifications for all sequences in the SH. An automatically or manually designated sequence is chosen to represent each such SH. These sequences are released (https://unite.ut.ee/repository.php) for use by the scientific community in, for example, local sequence similarity searches and next-generation sequencing analysis pipelines. The system and the data are updated automatically as the number of public fungal ITS sequences grows.
UNITE 提供了一套统一的方法,用于界定、识别、传播以及处理基于DNA的物种假说(Species Hypotheses,SH)。通过应用一套动态距离阈值(0.5%~3.0%),国际核酸序列数据库中的所有真菌内转录间隔区(Internal Transcribed Spacer,ITS)序列可被聚类至近似物种水平的聚类簇。所有物种假说(SH)均会被赋予一个以数字对象标识符(Digital Object Identifier,DOI)形式存在的唯一且稳定的名称,其分类学与生态学注释信息将通过分布式、基于网页的第三方注释工作完成核验。物种假说(SH)会尽可能结合该聚类簇内所有序列的鉴定结果,与对应分类单元名称及其分类层级(如门、纲、目等)建立关联。研究人员将选取一条自动或手动指定的序列作为每个此类物种假说(SH)的代表序列。这些代表序列已发布于(https://unite.ut.ee/repository.php),可供科研界使用,例如用于本地序列相似性检索以及高通量测序分析流程中。随着公共真菌ITS序列数量的持续增长,该系统与配套数据集将自动进行更新。



