DSGA components of various TCGA datasets
收藏资源简介:
This dataset provides processed bulk expression data components derived from multiple TCGA cancer cohorts. The TCGA expression datasets were obtained from the UCSC Xena platform. The raw data underwent a standardized pre-processing pipeline which included:Data pre-processing, Normalization & Transformation using the `voom` function from the `limma` R package. Following this pipeline, the DSGA algorithm (Nicolau, Monica, et al. "Disease-specific genomic analysis: identifying the signature of pathologic biology." Bioinformatics 23.8 (2007): 957-965. ) was applied to the processed expression data to generate the resulting data components. This record contains:* The primary DSGA-derived component files for each TCGA dataset analyzed.* The corresponding clinical survival data (e.g., Overall Survival time and status) for the patients associated with each cohort. This processed data is intended for researchers interested in analyzing TCGA data using the DSGA framework, validating findings, or for downstream analyses utilizing these pre-computed components.



