遇见数据集

Pan-mammalian pan-cancer RNA-seq datasets processed with Paipu

收藏
Zenodo2026-05-14 更新2026-05-26 收录
官方服务:

资源简介:

This dataset is associated with the manuscript "The Paipu framework enables creation of a large-scale mammalian cancer transcriptomics atlas". It contains pan-mammalian pan-cancer RNA-seq data processed using the Paipu pipeline. This data was compiled from NCBI SRA and processed consistently to generate gene expression data and harmonized sample metadata. There are two versions of the dataset: With duplicates: includes all samples after processing (Paipu_with_dups_*.tsv) Deduplicated: highly similar samples removed using sample expression correlation within species (Paipu_deduplicated_*.tsv) Each dataset includes sample metadata (*_metadata.tsv) and a gene expression matrix (*_expression.tsv). Gene expression values were normalized using trimmed mean of M-values (TMM) and transformed to log2 counts per million (CPM) with a prior count of 1.

提供机构:
Zenodo
创建时间:
2026-05-14
二维码
社区交流群
二维码
科研交流群
商业服务