Genome collection of gut bacterium Ruminococcus gnavus
收藏资源简介:
Collection of high-quality whole-genome assemblies of the gut bacterium Ruminococcus gnavus. Sequences are included as gzipped fasta files in sequences/ and functional annotations as generated by Bakta are included as well (gene_annotations/; GenBank format, GFF3, TSV and TXT). This directory also contains Panaroo files that were used in the GWAS analysis as described in the manuscript (subdirectory gene_annotations/panaroo). For isolate genomes sequenced using PacBio Circular Consensus Sequencing, methylation data are available under methylation/. There is a metadata file describing the origin of samples, when known. With this file, genome sequences can be related back to for example the age, sex and country of the person from whom the isolate was obtained. Finally, there is a file with md5 checksums that may be used to verify file integrity. This dataset was generated for the project described in this preprint: https://doi.org/10.1101/2024.06.27.600998 The manuscript has been published in Nature Communications: https://doi.org/10.1038/s41467-025-56449-x Changelog: Version 5.1: - includes panaroo file pan_genome_reference.fa with the nucleotide sequences of all the genes Version 5: - includes panaroo pangenome analysis files, which may be used to study the candidate genes identified in the GWAS Version 4.1: - has an updated metadata file including ENA accession numbers for the assemblies generated in this project. Version 4.1 has the public 'GCA' accession numbers for assemblies, rather than the internal 'ERZ' numbers which were in version 4.



