Rawdata for:Improving environmental DNA-based biodiversity monitoring through manual taxonomic curation in a tropical river in Colombia
收藏资源简介:
This dataset contains raw high-throughput sequencing data generated for an environmental DNA (eDNA) metabarcoding study of vertebrate diversity in the Magdalena River (Colombia). The files correspond to the direct output of Illumina sequencing runs and have not been demultiplexed or processed. Sequencing was performed using two primer sets targeting vertebrates and fishes: Vert01 and Teleo. A total of four sequencing runs are included, identified by the suffix “AX” in the file names (e.g., A1, A2, A5, A7). Runs A1 and A5 correspond to the Vert01 primer, while runs A2 and A7 correspond to the Teleo primer. Each sequencing run contains paired-end reads in compressed FASTQ format: *_R1.fastq.gz: forward reads *_R2.fastq.gz: reverse reads File names follow the Illumina naming convention (e.g., VL324___MB0423A2___R1.fastq.gz), where the “AX” identifier denotes the sequencing run. These files represent raw sequencing output prior to any bioinformatic processing, including demultiplexing, primer trimming, quality filtering, denoising, or taxonomic assignment. All scripts used for data processing, including demultiplexing, primer trimming, and downstream analyses (e.g., DADA2 workflow), are available in the next GitHub repository (https://github.com/jorgemorenoat-cell/Magdalena-Project-eDNA-Metabarcoding-Study-Moreno-Tilano-et-al..git). This repository is linked to the present dataset to ensure full reproducibility of the analyses



