官方服务:
资源简介:
Table for translating accession IDs into taxon names.
应用场景:
创建时间:
2018-01-01
相关数据集
Signor_3class_clustered-50
该数据集包含结构化序列数据,主要特征包括IdA(字符串)、IdB(字符串)、SeqA(字符串)、SeqB(字符串)和labels(整型)。数据集划分为训练集(12,491个样本)、测试集(1,018个样本)和验证集(1,061个样本),总大小约20.7MB。数据文件按默认配置存储在train-*、test-*和valid-*路径下。未提供具体任务描述,但标签字段暗示可能用于分类或匹配任务。
Hugging Face2026-02-12 更新120
Categories and taxa for sequence aggregation.
‘food’, ‘parasite’, ‘contaminant’, ‘unicellular eukaryotes’). The taxonomic levels to which groups of related sequences were aggregated within these categories are given as well as their common names.
Figshare2015-12-02 更新60
Silva SSU r138.1 training set for IDTAXA
Training set for Silva 138.1 for use with the IDTAXA taxonomy caller within the R DECIPHER packageCitation for DECIPHER: Wright ES (2016). “Using DECIPHER v2.0 to Analyze Big Biological Sequence Data
DataCite Commons2024-03-01 更新40
Benchmark datasets and Script for DeepEukFinder
batch_predict_def.sh: script to get results from the test files, usage: (after environment is set to have DeepEukFinder in PATH) bash batch_predict_def.sh.test-data.zip: 20 benchmark datasets for Deep
Figshare2023-09-03 更新60
SILVA-to-NCBI mapping
Contains a mapping of Genbank accession numbers to NCBI taxon ids and names. The accessions are those that are the reference sequences for SILVA clusters.
DataONE2017-06-16 更新50



