遇见数据集

CaBuAr P1

收藏
Zenodo2023-03-02 更新2026-05-26 收录
数据链接:
官方服务:

资源简介:

<pre># California Burned Areas Dataset This is the first part of the dataset. ### Dataset Summary This dataset contains images from Sentinel-2 satellites taken before and after a wildfire. The ground truth masks are provided by the California Department of Forestry and Fire Protection and they are mapped on the images. ### Supported Tasks The dataset is designed to do binary semantic segmentation of burned vs unburned areas. ## Dataset Structure ### Dataset opening Dataset was compressed using `h5py` and BZip2 from `hdf5plugin`. **WARNING: `hdf5plugin` is necessary to extract data** ### Data Instances Each matrix has a shape of 5490x5490xC, where C is 12 for pre-fire and post-fire images, while it is 0 for binary masks. ### Data Fields In each HDF5 file, you can find post-fire, pre-fire images and binary masks. The file is structured in this way: ```bash ├── foldn │ ├── uid0 │ │ ├── pre_fire │ │ ├── post_fire │ │ ├── mask │ ├── uid1 │ ├── post_fire │ ├── mask │ ├── foldm ├── uid2 │ ├── post_fire │ ├── mask ├── uid3 ├── pre_fire ├── post_fire ├── mask ... ``` where `foldn` and `foldm` are fold names and `uidn` is a unique identifier for the wilfire. ### Data Splits There are 5 random splits whose names are: 0, 1, 2, 3 and 4. ## Dataset Creation ### Source Data #### Initial Data Collection and Normalization Data are collected directly from Copernicus Open Access Hub through the API. The band files are aggregated into one single matrix. </pre>

# 加州烧毁区域数据集 本数据集为该数据集的第一部分。 ### 数据集概述 本数据集包含哨兵-2号(Sentinel-2)卫星拍摄的野火发生前后的遥感影像。真实标签掩码(ground truth masks)由加州林业与消防局(California Department of Forestry and Fire Protection)提供,并已配准至对应影像中。 ### 支持任务 本数据集旨在针对烧毁区域与未烧毁区域开展二元语义分割任务。 ## 数据集结构 ### 数据集读取方式 本数据集通过`h5py`与`hdf5plugin`提供的BZip2算法完成压缩。**注意:需安装`hdf5plugin`方可解压数据** ### 数据实例 每个数据矩阵的尺寸为5490×5490×C:对于野火前后影像,C=12;对于二元掩码,C=0。 ### 数据字段 每个HDF5文件中包含野火后影像、野火前影像以及二元掩码。文件结构如下: bash ├── foldn │ ├── uid0 │ │ ├── pre_fire │ │ ├── post_fire │ │ ├── mask │ ├── uid1 │ ├── post_fire │ ├── mask │ ├── foldm ├── uid2 │ ├── post_fire │ ├── mask ├── uid3 ├── pre_fire ├── post_fire ├── mask ... 其中`foldn`与`foldm`为文件夹名称,`uidn`为对应野火的唯一标识符。 ### 数据划分 本数据集共包含5组随机划分的子集,其编号分别为0、1、2、3、4。 ## 数据集构建 ### 源数据 #### 初始数据收集与归一化 数据通过API直接从哥白尼开放获取中心(Copernicus Open Access Hub)获取,所有波段文件将被整合为单个数据矩阵。

提供机构:
Zenodo
创建时间:
2023-03-02
二维码
社区交流群
二维码
科研交流群
商业服务