遇见数据集

Summary statistics of long-read RNA-Seq data.

收藏
Figshare2022-11-04 更新2026-04-28 收录
官方服务:

资源简介:

Long read RNA-Seq data were obtained from four studies as indicated in row 1: Lee et al [22], Yang et al [21], this study, and Chappell et al [15]. The data were obtained from independent biological samples of Plasmodium falciparum indicated in row 2. Data from samples 1–10 were used to construct the transcriptomic catalog. Details of developmental stage and RNA fraction for each sample are shown in rows 3 and 4, respectively. Long-read sequencing platform is indicated in row 5. Processed read datasets obtained from each sample are indicated in row 6; datasets from PacBio sequencing have “lq” or “hq” suffixes to indicate low quality or high-quality assignment from Iso-Seq3 read processing. The dataset labels shown in this row are consistent with dataset source labels in S2 Table (columns 33, 36−38). The numbers of processed reads available for alignment in each dataset are shown in row 7. Row 8 shows the number of reads aligned to the P. falciparum 3D7 v3.2 reference genome [14] by minimap2 [84]. Row 9 shows the number of candidate non-redundant isoforms assigned by TAMA collapse [27] from the aligned reads for each dataset. (XLSX)

本数据集所使用的长读长RNA测序(long read RNA-Seq)数据来源于第1行标注的四项研究:Lee等人[22]、Yang等人[21]、本研究以及Chappell等人[15]。所有数据均取自第2行标注的恶性疟原虫(Plasmodium falciparum)独立生物学样本。样本1至10的测序数据被用于构建转录组目录。各样本的发育阶段及RNA组分细节分别列于第3行和第4行。长读长测序平台信息标注于第5行。各样本的经处理读段数据集标注于第6行;PacBio测序生成的数据集带有"lq"或"hq"后缀,分别代表Iso-Seq3读段处理流程中得到的低质量或高质量读段。本行所列的数据集标签与补充表S2(第33、36−38列)中的数据集来源标签保持一致。各数据集可用于比对的经处理读段数量列于第7行。第8行展示了通过minimap2工具[84]比对至恶性疟原虫3D7株v3.2版参考基因组[14]的读段数量。第9行展示了各数据集经TAMA collapse工具[27]从比对读段中鉴定得到的候选非冗余转录本亚型数量。(XLSX)

创建时间:
2022-11-04
二维码
社区交流群
二维码
科研交流群
商业服务