VLM-CASE_carla_dataset
收藏资源简介:
该数据集包含10,560个在CARLA中收集的前置摄像头帧(1280×720),涵盖三种路面条件(干燥、湿润、雪地)、三种天气类型(晴朗、大雨、浓雾)、白天/夜晚以及三种照明辅助级别,每个帧都标有四个场景上下文字段。数据集分为训练集(8,448帧)、验证集(1,056帧)和测试集(1,056帧),并按联合标签分布分层。
This dataset contains 10,560 front-facing camera frames (1280×720) collected in CARLA, covering three road surface conditions (dry, wet, snow-covered), three weather types (clear, heavy rain, dense fog), day/night scenarios, and three levels of lighting assistance. Each frame is annotated with four scene context fields. The dataset is split into a training set (8,448 frames), a validation set (1,056 frames), and a test set (1,056 frames), with stratification performed based on the joint label distribution.
数据集概述
数据集名称:VLM-CASE 场景上下文数据集(Scene-Context Dataset)
数据集规模:共 10,560 张前视摄像头帧图像,分辨率为 1280×720。
数据来源:使用 CARLA 仿真模拟器采集。
场景覆盖:图像覆盖三种路面(干燥、湿滑、积雪)、三种天气(晴朗、大雨、浓雾)、白天/夜晚以及三种照明辅助等级。
标注内容:每张图像标注了四个场景上下文字段。
数据集划分:
- 训练集:8,448 张
- 验证集:1,056 张
- 测试集:1,056 张
- 划分方式:基于联合标签分布进行分层抽样。
下载地址:Hugging Face 数据集页面 ytj254/VLM-CASE_carla_dataset
标签说明:详细标签模式见仓库中的 dataset/README.md 文件。
用途:该数据集用于微调视觉语言模型(VLM),使其能够根据前视摄像头画面推理驾驶场景,并参数化生成上下文自适应安全包络(CASE),进而用于模型预测控制器的安全自主驾驶。
许可协议:数据、图表和视频以 CC-BY-4.0 许可发布。



