Identification of protein domains is a key step for understanding protein function. Hidden Markov Models (HMMs) have proved to be a powerful tool for this task. The Pfam database notably provides a la
Depth of coverage calculation is an important and computationally intensive preprocessing step in a variety of next generation sequencing pipelines, including the analyses of RNA-seq data, detection o
Comparison of the richness (Chao1) and diversity (Inverse Simpson) indexes for clustering-first pipelines before taxonomic merging, on the 200(V3) HC dataset with sequencing errors simulation when gen
The length of the screening window surrounding predicted breakpoints is 400 bp. “GS-ov” SVs are referred to as gold-standard overlapping SVs. “l/r match” represents SVs with partial-boundary-match (le
Additional file 7. Percentage of sequence coverage modelled into the exemplary mmCIF/PDBx structures used to optimize the COSNeti workflow. (A) The yeast (Saccharomyces cerevisiae PDB ID: 6snt), (B) r