Custom and Standard reference protein datasets used for publication Computational challenges to reproducibility, robustness, extensibility and reuse in multi-omics: a meta-workflow-based case study
收藏资源简介:
Customized protein sequence databases used in the study: Computational challenges to reproducibility, robustness, extensibility and reuse in multi-omics: a meta-workflow-based case study. Custom datasets: custom_merged_OV.fasta, custom_merged_BRCA.fasta, custom_merged_CRC95.fasta Reference datasets: hg19_refseq_pro_seq.fasta, humanRefSeq_Version54_with_tryp_DECOY.fasta These datasets have been used with the Docker image archived in http://doi.org/10.5281/zenodo.4612080
本研究使用的定制化蛋白质序列数据库:多组学(multi-omics)研究中可重复性、稳健性、可扩展性与数据复用性面临的计算挑战——一项基于元工作流(meta-workflow)的案例研究。定制数据集包括:custom_merged_OV.fasta、custom_merged_BRCA.fasta、custom_merged_CRC95.fasta。参考数据集包括:hg19_refseq_pro_seq.fasta、humanRefSeq_Version54_with_tryp_DECOY.fasta。上述数据集已与存档于http://doi.org/10.5281/zenodo.4612080的Docker镜像配合使用。



