Pre-ranking Synthetic Dataset for Spatial Prioritization of Environmental Monitoring Sites in Marajó Bay
收藏资源简介:
This dataset contains the pre-ranking synthetic dataset used in a methodological framework for the spatial prioritization of environmental monitoring sites in Marajó Bay, Brazilian Amazon. The table stores candidate sites, reference localities, spatial distances, auxiliary modeling factors, synthetic indicators, aggregated criteria, global risk scores, and logistical cost variables used before multicriteria ranking and budget-constrained optimization. Synthetic indicators were generated to support a controlled, transparent, and reproducible analytical scenario in a context with limited homogeneous in situ observations. Data Dictionary Field Data type Description ID string Unique identifier of the candidate site. Referencia string Reference locality name associated with the candidate site. Tipo_Ponto string General point type recorded in the dataset. Perfil_Cidade string Territorial profile of the reference locality. Sede_Lat float Latitude of the reference locality. Sede_Lon float Longitude of the reference locality. Lat float Latitude of the candidate site used in the analytical dataset. Lon float Longitude of the candidate site used in the analytical dataset. Localizacao string Text description of how the point was positioned relative to the water body. Massa_Dagua string Water body associated with the candidate site. Municipio_Shape string Municipality or municipalities associated with the spatial unit in the source geography. Dist_Sede_MassaDagua_km float Distance, in kilometers, between the reference locality and the adjusted point on the water body. Custo_Km_R$ float Cost per kilometer in Brazilian reais used in the logistics cost calculation. Municipio_Ancora string Anchor municipality or anchor locality associated with the point. Perfil_Ancora string Territorial profile of the anchor locality. Dist_Borda_Estuario_km float Distance, in kilometers, from the candidate site to the estuarine edge. Dist_Cidade_Proxima_km float Distance, in kilometers, from the candidate site to the nearest city or reference locality. Influencia_Cidade float City influence index stored in the pre-ranking dataset. The manuscript indicates that proximity to cities is used in the synthetic data generation process. Influencia_Borda float Edge influence index stored in the pre-ranking dataset. The manuscript indicates that proximity to the estuarine edge is used in the synthetic data generation process. Penalidade_Agua_Aberta float Open-water penalty term stored in the pre-ranking dataset. Fator_Ancoragem_Cidade float Auxiliary city-anchoring factor stored in the pre-ranking dataset. Fator_Ausencia_Dados float Auxiliary missing-data factor stored in the pre-ranking dataset. Fator_Prioridade_Espacial float Auxiliary spatial-priority factor stored in the pre-ranking dataset. IES float Sewage/Sanitary Sewerage Index. Higher values represent better sewerage conditions and therefore lower socio-environmental risk. ICR float Regular Solid Waste Collection Indicator. Higher values represent better waste collection conditions and therefore lower risk. IAA float Water Supply Index. Higher values represent better formal water supply conditions and therefore lower risk. IDEM float Demographic Density indicator. Higher values indicate greater human occupation pressure. IURB float Urbanization Index. Higher values indicate stronger urban pressure in the site surroundings. IUTU float Urban Land Use indicator. Higher values indicate greater urban land-use pressure. IAGR float Agricultural Land Use indicator. Higher values indicate greater agricultural pressure in the surroundings. IPAS float Pasture indicator. Higher values indicate greater pasture-related land occupation. IVEG float Natural Vegetation indicator. Higher values represent greater natural vegetation cover and therefore lower risk. DRSAI float Diseases Related to Inadequate Environmental Sanitation. Higher values indicate greater socio-environmental criticality. IVA float Environmental Vulnerability Index. Higher values indicate greater environmental vulnerability. C1 float Main criterion 1: basic sanitation, derived from IES after transformation to the model’s risk scale. C2 float Main criterion 2: solid waste, derived from ICR after transformation to the model’s risk scale. C3 float Main criterion 3: water supply, derived from IAA after transformation to the model’s risk scale. C4 float Main criterion 4: anthropogenic pressure, aggregated from IDEM and IURB. C5 float Main criterion 5: land use and land cover, aggregated from IUTU, IAGR, IPAS, and IVEG. C6 float Main criterion 6: socio-environmental vulnerability, aggregated from DRSAI and IVA. RISCO_GLOBAL float Global risk score computed as a weighted synthetic measure of criticality across indicator groups. Qtde_Indicadores_Ausentes integer Number of missing indicators in the record. Indicadores_Ausentes string List of missing indicators in the record, when applicable. Qtde_Criterios_Ausentes integer Number of missing aggregated criteria in the record. Criterios_Ausentes string List of missing aggregated criteria in the record, when applicable. Distancia_Logistica_km float Estimated logistical distance to access the candidate site, in kilometers. Tempo_Logistico_h float Estimated logistical travel time to access the candidate site, in hours. Metodo_Custo_Logistico string Method used to estimate logistical cost. Custo_R$ float Final estimated logistical cost for the candidate site, in Brazilian reais.



