Binned Nanopore/Illumina Reads From DNAformer, Split by the Data Type
收藏资源简介:
This dataset includes binned DNA sequencing reads from Nanopore and Illumina platforms, used in DNA data storage experiments (DNAformer). Each file contains read clusters (label, separator of "****", reads) separated by two blank lines. Files are split based on the file type: _Random.txt: clusters of the random file. _Semantic.txt: clusters of the semantic file. Includes data from two Nanopore flowcells (separate and merged) and a test Illumina set. Additionally, the semantic file and the random file are included in the repository as well.
本数据集包含源自纳米孔(Nanopore)与因美纳(Illumina)测序平台的分箱DNA测序读段,用于DNA数据存储实验(DNAformer)。每个文件包含以两条空行分隔的测序读段簇(由标签、分隔符"****"与读段组成)。 数据集按文件类型划分为: _Random.txt:随机文件对应的测序读段簇。 _Semantic.txt:语义文件对应的测序读段簇。 本数据集包含来自两个纳米孔测序流动槽(独立数据集与合并数据集)以及一组因美纳测序测试集的数据。 此外,随机文件与语义文件也已一并收录至存储库中。



