TCRalign benchmark dataset, case study dataset, and precomputed MSAs
收藏资源简介:
SPTC_structures.7z – Benchmark dataset for calculating SP and TC scores. strike_1K_512K_dataset.7z – Benchmark dataset for calculating strike scores, as well as for memory and runtime testing. case_dataset.tar.gz – Case study dataset, including TCR sequences from three cohorts. tcr_alignments.7z – Precomputed multiple sequence alignments (MSAs): hmm_output.sto – MSAs specifying template PDB IDs for structure prediction. single.a3m – MSAs including additional TCR sequences obtained from the iReceptor database. uniprot_hits.sto – MSAs originally generated by OpenFold and subsequently realigned using TCRalign.
SPTC_structures.7z – 用于计算SP与TC得分的基准数据集。 strike_1K_512K_dataset.7z – 用于计算strike得分,同时可用于内存与运行时测试的基准数据集。 case_dataset.tar.gz – 案例研究数据集,包含来自三个队列的TCR (T Cell Receptor) 序列。 tcr_alignments.7z – 预计算多序列比对(multiple sequence alignments, MSA)文件: hmm_output.sto – 用于指定结构预测所用模板PDB (Protein Data Bank) ID的多序列比对文件。 single.a3m – 包含从iReceptor数据库获取的额外TCR序列的多序列比对文件。 uniprot_hits.sto – 最初由OpenFold生成,随后通过TCRalign进行重比对的多序列比对文件。



