遇见数据集

An Arabidopsis gene expression matrix derived from large-scaled RNA-Seq datasets

收藏
NIAID Data Ecosystem2026-03-12 收录
官方服务:

资源简介:

The dataset contains a gene expression matrix to be used with the EXPLICIT package to construct an gene expression predictor. It has 24545 RNA-Seq samples with 38194 genes in total. Two versions of the matrix are provided. It is recommended to use the smaller version first (At.matrix.demo.h5) to get familiar with the EXPLICIT package, and then change to the full version (At.matrix.full.h5). The smaller matrix contains 5000 samples randomly selected from all 24545 samples. At.matrix.demo.h5 |--expression_log2cpm -------- (5000 samples [row] X 38194 genes [col]) |--gene_name ------------------ ( for the 38194 genes) |--rnaseq_id ------------------- (for the 5000 samples) |--idx_tf_gene ------- (specifying TF genes used for model construction) |--idx_target_gene --- (specifying target genes used for model construction) |--independent_samples_for_validation >>>|--expression_log2cpm ------ (2 samples [row] X 38194 genes [col]) >>>|--gene_name ----------------- (for the 38194 genes) >>>|--sample_id ---------------- (for the 2 RNA-Seq samples) At.matrix.full.h5 |--expression_log2cpm -------- (24545 samples [row] X 38194 genes [col]) |--gene_name ------------------ (for the 38194 genes) |--rnaseq_id ------------------- (for the 24545 samples) |--idx_tf_gene ------- (specifying TF genes used for model construction) |--idx_target_gene --- (specifying target genes used for model construction) |--independent_samples_for_validation >>>|--expression_log2cpm ------ (2 samples [row] X 38194 genes [col]) >>>|--gene_name ----------------- (for the 38194 genes) >>>|--sample_id ---------------- (for the 2 RNA-Seq samples)

创建时间:
2021-05-06
二维码
社区交流群
二维码
科研交流群
商业服务