遇见数据集

DNA methylation & adversity: My Body My Story (MBMS) epigenetic analytic dataset (R01 MD014304, 2019–2026)

收藏
DataONE2026-04-06 更新2026-05-19 收录
官方服务:

资源简介:

The DNA methylation & adversity study (NIH R01MD014304; Project Period: 09/01/2019–03/31/2026) investigates how DNA methylation (DNAm) varies with exposure to racial discrimination, economic hardship, and air pollution, and how these epigenetic changes may contribute to inequities in cardiometabolic disease risk and accelerated aging (epigenetic age > chronological age). This dataset record documents the My Body My Story (MBMS) epigenetic analytic dataset used for the R01MD014304 analyses. MBMS (NIH R01AG027122; Project Period: 9/15/2007–6/30/2013) is a cross‑sectional, population‑based study of racial discrimination and cardiometabolic health that recruited a random sample of 1,005 U.S‑born non‑Hispanic Black (n = 504) and non‑Hispanic White (n = 501) adults ages 35–64 from four community health centers in Boston, Massachusetts (2008–2010). MBMS includes detailed measures of racial discrimination and socioeconomic position across the life course, residential air pollution, cardiometabolic outcomes, and relevant covariates. For R01MD014304, stored MBMS dried blood spots were assayed in 2020–2021 using the Illumina MethylationEPIC BeadChip, generating genome‑wide DNAm data for 293 participants. The de‑identified epigenetic analytic dataset described here includes DNAm measures (including multiple epigenetic clocks and other DNAm‑based indices) linked to MBMS social, clinical, and contextual variables needed for the project’s main analyses. For comparison purposes, the R01MD014304 project also uses analogous data from the Multi‑Ethnic Study of Atherosclerosis (MESA) and publicly available U.S. Census and environmental data. These external datasets are not included in the MBMS epigenetic analytic dataset documented here; investigators must obtain them directly from the original sources. MESA data must be requested directly from the MESA study. Materials in this Dataverse record: This record provides formal citation information and the core materials needed to construct the de‑identified MBMS epigenetic analytic dataset. These materials are bundled in the file “Create-Analytic-MBMS-Dataset-main.zip” and mirrored in a dedicated GitHub repository for analytic dataset construction. Contents include: a data dictionary and codebook describing all variables in the analytic dataset; documentation of variable derivations and data sources; copies of MBMS survey and participant booklet instruments used in this project; and R scripts that create the analytic dataset described here. Neither this Dataverse record nor its GitHub mirror contains any individual‑level MBMS or MESA data. Additional public code and contextual data (hosted externally): Code used for specific published analyses and for generating public, non‑human‑subjects contextual measures is available in separate repositories. These repositories do not include MBMS or MESA individual‑level data but document analytic workflows and public inputs. Key repositories include: Code to generate multiple epigenetic clocks for MBMS and MESA datasets: epigenetic_clocks_mbms_mesa Code for the epigenome‑wide analysis (EWAS) of DNAm, racialized and economic inequities, and air pollution: EWAS_of_social_inequities DNAmAndAdversity_public R package and shared analysis functions (with private data omitted) (archived at Zenodo) Census‑derived ICE metrics and other contextual measures for MBMS and MESA participants, created from publicly available U.S. Census and American Community Survey data: CensusData A consolidated, up‑to‑date list of these and related resources is also maintained on the Krieger Research Group data sharing resources page. Data Storage: The de‑identified analytic dataset is stored on secure Harvard T.H. Chan School of Public Health servers managed by Harvard Chan School Information Technology. Data Access: Because the data contain sensitive health, genetic, and social information, the underlying analytic data files are not openly available for download from this repository. Researchers interested in accessing the MBMS epigenetic analytic dataset must apply for this access, via a governed access process and Data Use Agreement (DUA). To apply for data access, please complete the MBMS DNA Methylation Project application form. For details on eligibility, procedures, and data use conditions, see the PDF “EpiMBMS_ANALYTIC_DATASET_ACCESS+TERMS_OF_USE.pdf” included in this record.

本数据集对应**DNA甲基化与逆境研究(DNA methylation & adversity study,NIH R01MD014304;项目周期:2019年9月1日–2026年3月31日)**,旨在探究DNA甲基化(DNA methylation,DNAm)随种族歧视、经济困境与空气污染暴露的变化规律,以及这些表观遗传改变如何加剧心血管代谢疾病风险与加速衰老(表观遗传年龄>实足年龄)的健康不平等问题。本数据集记录面向该R01MD014304项目分析所用的“我的身体,我的故事(My Body My Story, MBMS)”表观遗传分析数据集。 MBMS(NIH R01AG027122;项目周期:2007年9月15日–2013年6月30日)是一项针对种族歧视与心血管代谢健康的横断面基于人群研究,于2008–2010年从马萨诸塞州波士顿市的4家社区卫生中心招募了1005名美国出生的35–64岁成年人,其中非西班牙裔黑人504名、非西班牙裔白人501名。MBMS收集了覆盖全生命周期的种族歧视、社会经济地位、居住空气污染、心血管代谢结局及相关协变量的详细测量数据。 针对本R01MD014304项目,研究团队于2020–2021年对存储的MBMS干血斑(dried blood spots)采用Illumina MethylationEPIC BeadChip芯片进行检测,为293名受试者生成了全基因组DNA甲基化数据。本记录所描述的去标识化表观遗传分析数据集,包含DNA甲基化测量指标(含多种表观遗传时钟及其他基于DNA甲基化的指数),并与本项目核心分析所需的MBMS社会、临床及环境变量进行了关联。 为便于对比,R01MD014304项目还使用了来自多种族动脉粥样硬化研究(Multi-Ethnic Study of Atherosclerosis, MESA)的同类数据,以及公开可得的美国人口普查与环境数据。本记录收录的MBMS表观遗传分析数据集未包含上述外部数据集,研究者需直接从原始来源获取。其中MESA数据需直接向MESA研究团队申请获取。 本Dataverse记录包含的材料:本记录提供了构建去标识化MBMS表观遗传分析数据集所需的正式引用信息与核心材料。这些材料打包于文件"Create-Analytic-MBMS-Dataset-main.zip"中,并在专属GitHub仓库中同步,用于分析数据集的构建。内容包括:分析数据集所有变量的数据字典与编码手册、变量推导及数据来源说明、本项目所用的MBMS问卷与受试者手册副本,以及生成本记录所述分析数据集的R脚本。本Dataverse记录及其GitHub镜像均未包含任何个体层面的MBMS或MESA数据。 额外的公开代码与环境数据(托管于外部平台):用于已发表的特定分析以及生成公开非人类受试者环境测量指标的代码,存储于独立仓库中。这些仓库未包含MBMS或MESA的个体层面数据,但记录了分析流程与公开输入数据。主要仓库包括: 1. 用于为MBMS与MESA数据集生成多种表观遗传时钟的代码:epigenetic_clocks_mbms_mesa 2. 针对DNA甲基化、种族与经济不平等及空气污染的表观全基因组关联分析(epigenome-wide association study, EWAS)代码:EWAS_of_social_inequities 3. DNAmAndAdversity_public R包与共享分析函数(已移除私有数据,存档于Zenodo) 4. 基于公开美国人口普查与美国社区调查数据生成的、针对MBMS与MESA受试者的人口普查衍生ICE指标及其他环境测量数据:CensusData 本研究团队还在克里格研究组(Krieger Research Group)的数据共享资源页面维护了上述及相关资源的整合最新列表。 数据存储:本去标识化分析数据集存储于哈佛大学陈曾熙公共卫生学院(Harvard T.H. Chan School of Public Health)由学院信息技术部门管理的安全服务器中。 数据获取:由于本数据集包含敏感的健康、遗传与社会信息,底层分析数据文件无法从本仓库直接公开下载。有意向获取MBMS表观遗传分析数据集的研究者需通过受控访问流程与数据使用协议(Data Use Agreement, DUA)申请访问权限。如需申请数据访问,请填写MBMS DNA甲基化项目申请表。关于资格要求、申请流程与数据使用条件的详细信息,请参阅本记录中包含的PDF文件"EpiMBMS_ANALYTIC_DATASET_ACCESS+TERMS_OF_USE.pdf"。

创建时间:
2026-04-09
二维码
社区交流群
二维码
科研交流群
商业服务