遇见数据集

Supplementary Datasets for dadasnake workflow

收藏
Zenodo2020-10-18 更新2026-05-25 收录
数据链接:
官方服务:

资源简介:

This dataset contains configuration and results files for the proof-of-principle of the dadasnake pipeline. Includes tables with the composition of ground-truth data or mock-communities. dadasnake wraps pre-processing of sequencing reads, delineation of exact sequencing variants using the favorably benchmarked, widely-used the DADA2 algorithm, taxonomic classification and post-processing of the resultant tables, and hand-off in standard formats, into a user-friendly, one-command Snakemake pipeline. The suitability of the provided default configurations is demonstrated using mock-community data from bacteria and archaea, as well as fungi. By use of Snakemake, dadasnake makes efficient use of high-performance computing infrastructures. Easy user configuration guarantees flexibility of all steps, including the processing of data from multiple sequencing platforms. dadasnake facilitates easy installation via conda environments. dadasnake is available at https://github.com/a-h-b/dadasnake .

本数据集包含dadasnake流程原理验证所用的配置文件与结果文件,其中涵盖基准真值(ground-truth)数据或模拟群落(mock-communities)的组成表格。dadasnake将测序读段预处理、使用经过充分基准测试且广泛应用的DADA2算法划定精确测序变异体、分类学分类与结果表格后处理,以及以标准格式进行数据移交等步骤,封装为易用的单命令Snakemake工作流。所提供的默认配置的适用性,通过细菌、古菌与真菌的模拟群落数据得到了验证。借助Snakemake,dadasnake可高效利用高性能计算基础设施。便捷的用户配置可保障所有步骤的灵活性,支持处理来自多种测序平台的数据。dadasnake可通过Conda环境实现便捷安装。dadasnake的开源仓库地址为https://github.com/a-h-b/dadasnake。

提供机构:
Zenodo
创建时间:
2020-10-18
二维码
社区交流群
二维码
科研交流群
商业服务