遇见数据集

Q Score Segmented FAST5 Evaluation Data Set

收藏
Zenodo2024-06-21 更新2026-05-26 收录
官方服务:

资源简介:

Q Score segmented raw FAST5 data set used for accuracy characterization of the novel Alignment Matrix soft decoding algorithm (https://doi.org/10.5281/zenodo.11454877) applied to the HEDGES DNA-information storage code. Implementation of the HEDGES code used for accuracy assessment of our algorithm is based on the publication of Press et al. (https://doi.org/10.1073/pnas.2004821117). Each archive in this data set generally corresponds to a certain design length and HEDGES rate. For example, 1250, 1667, 3333, and 5000 correspond to HEDGES rates of 0.125, 0.167, 0.33, and 0.5 respectively. Additionally, archives labeled with "half" and "quarter" indicate DNA molecule designs that are approximately half and quarter the length of archives labeled "full". Archives labeled with "s1" or "s2" correpsond to data for strands indexed as 1 and 2 for the 0.167 hedges "full" design. Within each archive are FAST5 directories that each correspond to Q Score segment ranges that were used to evaluate the impact of Q Score on soft decoding byte error rate. Each FAST5 directory is clearly labeled with the start and end Q Score value that was used to construct the data set.

提供机构:
Zenodo
创建时间:
2024-06-18
二维码
社区交流群
二维码
科研交流群
商业服务