Popgen Stats from Hufford et al. 2012 Nat. Gen.
收藏资源简介:
Summary statistics for 10kb windows genome-wide and for genes in the maize v2 filtered gene set. Data are from Hufford et al. 2012: http://www.nature.com/ng/journal/v44/n7/full/ng.2309.html See details in the paper for criteria for calling SNPs, data used for statistics, etc. Columns are: locus: GRM name of gene in the filtered gene set S: number of Segregating sites ThetaW: Watterson's estimate of theta (per locus) ThetaPi: nucleotide diversity (per locus) ThetaH: Fay and Wu (2000) estimator (per locus) TajD: Tajima's D seqbp: # of bp sequenced. this should be used as the denominator to calculate per bp. values of the above statistics.
本数据集包含玉米v2过滤基因集内全基因组范围的10kb窗口区域及基因区域的汇总统计数据。数据来源于Hufford等(2012)的研究成果:http://www.nature.com/ng/journal/v44/n7/full/ng.2309.html。有关单核苷酸多态性(SNP, Single Nucleotide Polymorphism)的调用标准、统计分析所用数据集等详细信息,请参见该论文。 各列含义如下: locus:过滤基因集中基因的GRM名称 S:分离位点(Segregating sites)数量 ThetaW:沃特斯θ估计值(Watterson's estimate of theta,单基因座水平) ThetaPi:核苷酸多样性(nucleotide diversity,单基因座水平) ThetaH:Fay与Wu(2000)估计量(Fay and Wu (2000) estimator,单基因座水平) TajD:田岛D统计量(Tajima's D) seqbp:测序碱基对总数量,该值可作为分母用于计算上述各统计量的每碱基对数值。



