遇见数据集

whittle benchmark datasets: length-stratified HG002 ONT subsets

收藏
Zenodo2026-07-14 更新2026-08-02 收录
官方服务:

资源简介:

Length-stratified subsets of the public Oxford Nanopore HG002 dorado-SUP release, used to benchmark the whittle long-read trimmer (https://github.com/erdikilic/whittle). Two normalizations are provided: eqbase holds sequence volume constant at 180 Mb per subset (57,883 / 8,848 / 4,103 reads for short / mid / long), and eqread holds read count constant at 8,000 reads per subset. Each of the six subsets spans a distinct read-length regime (N50 approximately 5, 21, and 42 kb) and is provided both as unaligned BAM carrying MM/ML base-modification tags and as gzip-compressed FASTQ. Derived from ONT Open Data.

提供机构:
Zenodo
创建时间:
2026-07-14
二维码
社区交流群
二维码
科研交流群
商业服务