This fileset contains a written review of v2 of Heng Li's manuscript on bwa mem, as retrieved from arxiv.org. It also includes some basic benchmarking datasets and accuracy evaluation code applied to
We used both simulated and experimental datasets derived from human genomic DNA, human T cell receptor repertoires, and intra-host viral populations. Next, we summarize datasets shared here, i.e., D1,
Additional file 3: Table of Prider benchmark test files’ metadata and system.time output for the increasing number of sequences dataset. The data includes the number of sequences, the number of bases,
Annotations for each of the SV calls as well as likely non-SV regions from the PacBio aligned sequence dataset for NA12878 using svclassify. (CSV 1.88 kb)