遇见数据集

Pre-ranking Synthetic Dataset for Spatial Prioritization of Environmental Monitoring Sites in Marajó Bay

收藏
Zenodo2026-05-19 更新2026-05-26 收录
官方服务:

资源简介:

This dataset contains the pre-ranking synthetic dataset used in a methodological framework for the spatial prioritization of environmental monitoring sites in Marajó Bay, Brazilian Amazon. The table stores candidate sites, reference localities, spatial distances, auxiliary modeling factors, synthetic indicators, aggregated criteria, global risk scores, and logistical cost variables used before multicriteria ranking and budget-constrained optimization. Synthetic indicators were generated to support a controlled, transparent, and reproducible analytical scenario in a context with limited homogeneous in situ observations. Data Dictionary Field Data type Description ID string Unique identifier of the candidate site. Referencia string Reference locality name associated with the candidate site. Tipo_Ponto string General point type recorded in the dataset. Perfil_Cidade string Territorial profile of the reference locality. Sede_Lat float Latitude of the reference locality. Sede_Lon float Longitude of the reference locality. Lat float Latitude of the candidate site used in the analytical dataset. Lon float Longitude of the candidate site used in the analytical dataset. Localizacao string Text description of how the point was positioned relative to the water body. Massa_Dagua string Water body associated with the candidate site. Municipio_Shape string Municipality or municipalities associated with the spatial unit in the source geography. Dist_Sede_MassaDagua_km float Distance, in kilometers, between the reference locality and the adjusted point on the water body. Custo_Km_R$ float Cost per kilometer in Brazilian reais used in the logistics cost calculation. Municipio_Ancora string Anchor municipality or anchor locality associated with the point. Perfil_Ancora string Territorial profile of the anchor locality. Dist_Borda_Estuario_km float Distance, in kilometers, from the candidate site to the estuarine edge. Dist_Cidade_Proxima_km float Distance, in kilometers, from the candidate site to the nearest city or reference locality. Influencia_Cidade float City influence index stored in the pre-ranking dataset. The manuscript indicates that proximity to cities is used in the synthetic data generation process. Influencia_Borda float Edge influence index stored in the pre-ranking dataset. The manuscript indicates that proximity to the estuarine edge is used in the synthetic data generation process. Penalidade_Agua_Aberta float Open-water penalty term stored in the pre-ranking dataset. Fator_Ancoragem_Cidade float Auxiliary city-anchoring factor stored in the pre-ranking dataset. Fator_Ausencia_Dados float Auxiliary missing-data factor stored in the pre-ranking dataset. Fator_Prioridade_Espacial float Auxiliary spatial-priority factor stored in the pre-ranking dataset. IES float Sewage/Sanitary Sewerage Index. Higher values represent better sewerage conditions and therefore lower socio-environmental risk. ICR float Regular Solid Waste Collection Indicator. Higher values represent better waste collection conditions and therefore lower risk. IAA float Water Supply Index. Higher values represent better formal water supply conditions and therefore lower risk. IDEM float Demographic Density indicator. Higher values indicate greater human occupation pressure. IURB float Urbanization Index. Higher values indicate stronger urban pressure in the site surroundings. IUTU float Urban Land Use indicator. Higher values indicate greater urban land-use pressure. IAGR float Agricultural Land Use indicator. Higher values indicate greater agricultural pressure in the surroundings. IPAS float Pasture indicator. Higher values indicate greater pasture-related land occupation. IVEG float Natural Vegetation indicator. Higher values represent greater natural vegetation cover and therefore lower risk. DRSAI float Diseases Related to Inadequate Environmental Sanitation. Higher values indicate greater socio-environmental criticality. IVA float Environmental Vulnerability Index. Higher values indicate greater environmental vulnerability. C1 float Main criterion 1: basic sanitation, derived from IES after transformation to the model’s risk scale. C2 float Main criterion 2: solid waste, derived from ICR after transformation to the model’s risk scale. C3 float Main criterion 3: water supply, derived from IAA after transformation to the model’s risk scale. C4 float Main criterion 4: anthropogenic pressure, aggregated from IDEM and IURB. C5 float Main criterion 5: land use and land cover, aggregated from IUTU, IAGR, IPAS, and IVEG. C6 float Main criterion 6: socio-environmental vulnerability, aggregated from DRSAI and IVA. RISCO_GLOBAL float Global risk score computed as a weighted synthetic measure of criticality across indicator groups. Qtde_Indicadores_Ausentes integer Number of missing indicators in the record. Indicadores_Ausentes string List of missing indicators in the record, when applicable. Qtde_Criterios_Ausentes integer Number of missing aggregated criteria in the record. Criterios_Ausentes string List of missing aggregated criteria in the record, when applicable. Distancia_Logistica_km float Estimated logistical distance to access the candidate site, in kilometers. Tempo_Logistico_h float Estimated logistical travel time to access the candidate site, in hours. Metodo_Custo_Logistico string Method used to estimate logistical cost. Custo_R$ float Final estimated logistical cost for the candidate site, in Brazilian reais.

本数据集为巴西亚马逊地区的马拉若湾(Marajó Bay)环境监测站点空间优先级规划方法框架中所用的预排序合成数据集。该数据表存储了多准则排序与预算约束优化前使用的候选站点、参考点位、空间距离、辅助建模因子、合成指标、聚合准则、全局风险评分以及后勤成本变量。鉴于原位同质观测数据有限,本研究生成了合成指标,以支持具备可控性、透明性与可复现性的分析场景。 数据字典 字段 数据类型 字段说明 ID string 候选站点的唯一标识符。 Referencia string 与候选站点关联的参考点位名称。 Tipo_Ponto string 数据集中记录的通用点位类型。 Perfil_Cidade string 参考点位的区域概况。 Sede_Lat float 参考点位的纬度。 Sede_Lon float 参考点位的经度。 Lat float 分析数据集所用候选站点的纬度。 Lon float 分析数据集所用候选站点的经度。 Localizacao string 点位相对于水体的位置文本描述。 Massa_Dagua string 与候选站点关联的水体。 Municipio_Shape string 源地理空间单元关联的市镇。 Dist_Sede_MassaDagua_km float 参考点位与水体上调整后点位间的距离,单位:千米。 Custo_Km_R$ float 后勤成本计算中使用的每千米成本,单位:巴西雷亚尔。 Municipio_Ancora string 与点位关联的锚定市镇或锚定点位。 Perfil_Ancora string 锚定点位的区域概况。 Dist_Borda_Estuario_km float 候选站点至河口边缘的距离,单位:千米。 Dist_Cidade_Proxima_km float 候选站点至最近城市或参考点位的距离,单位:千米。 Influencia_Cidade float 预排序数据集中存储的城市影响力指数。相关研究手稿表明,合成数据生成过程中会用到点位与城市的邻近性信息。 Influencia_Borda float 预排序数据集中存储的边缘影响力指数。相关研究手稿表明,合成数据生成过程中会用到点位与河口边缘的邻近性信息。 Penalidade_Agua_Aberta float 预排序数据集中存储的开阔水域惩罚项。 Fator_Ancoragem_Cidade float 预排序数据集中存储的辅助城市锚定因子。 Fator_Ausencia_Dados float 预排序数据集中存储的辅助缺失数据因子。 Fator_Prioridade_Espacial float 预排序数据集中存储的辅助空间优先级因子。 IES float 污水处理/卫生下水道指数。数值越高代表下水道卫生条件越好,社会环境风险越低。 ICR float 常规固体废物收集指数。数值越高代表垃圾收集条件越好,相关风险越低。 IAA float 供水指数。数值越高代表正规供水条件越好,相关风险越低。 IDEM float 人口密度指数。数值越高代表人类活动占用压力越大。 IURB float 城市化指数。数值越高代表站点周边的城市压力越强。 IUTU float 城市土地利用指数。数值越高代表城市土地利用压力越大。 IAGR float 农业土地利用指数。数值越高代表周边农业活动压力越大。 IPAS float 牧场指数。数值越高代表与牧场相关的土地占用程度越高。 IVEG float 自然植被指数。数值越高代表自然植被覆盖度越高,相关风险越低。 DRSAI float 环境卫生不足相关疾病指数。数值越高代表社会环境临界性越强。 IVA float 环境脆弱性指数。数值越高代表环境脆弱性越强。 C1 float 主准则1:基础卫生设施,由IES转换至模型风险尺度后得到。 C2 float 主准则2:固体废物管理,由ICR转换至模型风险尺度后得到。 C3 float 主准则3:供水保障,由IAA转换至模型风险尺度后得到。 C4 float 主准则4:人为压力,由IDEM与IURB聚合得到。 C5 float 主准则5:土地利用与土地覆盖,由IUTU、IAGR、IPAS与IVEG聚合得到。 C6 float 主准则6:社会环境脆弱性,由DRSAI与IVA聚合得到。 RISCO_GLOBAL float 全局风险评分,通过对各指标组的临界性进行加权合成计算得到。 Qtde_Indicadores_Ausentes integer 该记录中缺失的指标数量。 Indicadores_Ausentes string 适用时,该记录中缺失的指标列表。 Qtde_Criterios_Ausentes integer 该记录中缺失的聚合准则数量。 Criterios_Ausentes string 适用时,该记录中缺失的聚合准则列表。 Distancia_Logistica_km float 访问候选站点的估算后勤距离,单位:千米。 Tempo_Logistico_h float 访问候选站点的估算后勤出行时间,单位:小时。 Metodo_Custo_Logistico string 用于估算后勤成本的方法。 Custo_R$ float 候选站点的最终估算后勤成本,单位:巴西雷亚尔。

提供机构:
Zenodo
创建时间:
2026-05-19
二维码
社区交流群
二维码
科研交流群
商业服务