VLA_Benchmark_Prompted_1200
收藏官方服务:
资源简介:
VLA Benchmark Prompted 1200 是一个非破坏性的LeRobot v3.0合并数据集,由六个各含200个episode的基准数据集整合而成。该数据集包含1,200个训练片段,总计615,537帧。每个片段的task string均来自prompt_manifest_v2.csv中指定的提示。在数据标注中,Coke和Pepsi对象被统一规范标注为压碎的Coke罐和压碎的Pepsi罐。所有数据行均属于训练分割,无测试或验证分割。
VLA Benchmark Prompted 1200 is a non-destructive LeRobot v3.0 merged dataset, integrating six benchmark datasets each containing 200 episodes. It comprises 1,200 training episodes with a total of 615,537 frames. The task string for each episode is derived from specified prompts in prompt_manifest_v2.csv. In the annotations, Coke and Pepsi objects are uniformly labeled as crushed Coke cans and crushed Pepsi cans. All data rows belong to the training split, with no test or validation splits.
提供机构:
justintiensmith创建时间:
2026-07-27
原始信息汇总
数据集概述
- 数据集名称: VLA Benchmark Prompted 1200
- 许可证: Apache-2.0
- 任务类型: 机器人学(Robotics)
- 标签: LeRobot, VLA, Robotics
- 数据集配置: 默认配置,数据文件位于
data/*/*.parquet
数据规模
- 训练集: 1,200 个训练片段(episodes),共 615,537 帧(frames)
- 所有行均属于训练集划分
数据来源与构成
- 该数据集是六个各包含 200 个片段的标准数据集的非破坏性 LeRobot v3.0 合并结果
- 六个源数据集及其不可变版本(commit hash)如下:
justintiensmith/MT_State_Recognition_200justintiensmith/SP_Relational_Placement_200justintiensmith/SP_Referential_Disambiguation_200justintiensmith/SP_Sequencing_200justintiensmith/SP_Counting_200justintiensmith/MT_Size_Recognition_200
任务与提示设计
- 每个片段的任务字符串来自
meta/benchmark/prompt_manifest_v2.csv中的指定提示 - 提示设计(Prompt Manifest v2)仅修改了 200 个状态识别任务字符串:
open和upside-down在微调时显式显示closed和upright在微调语言中被保留(hold out)closed/upright的物理演示仍保留在训练数据中,但仅使用类别提示(如Place a white cup in the bowl.)- 类别提示在所有四个物理状态/方向单元格中共享,通用措辞不代表
closed或upright
- 推荐采样器权重为审核元数据(manifest 中的 audit metadata),本数据集未重复或过采样片段
数据内容说明
- Coke 和 Pepsi 对象的标准标签为:压扁的 Coke 罐和压扁的 Pepsi 罐
- 视频和轨迹与提示清单 v1 版本相比保持不变
- 活跃的电子表格位于
meta/benchmark/prompt_manifest_v2.xlsx,v1 清单和验证文件保留以供溯源
本地标识符
justintiensmith/VLA_Benchmark_Prompted_1200



