遇见数据集

rollout_folding_smolvla_full_20260722_150412

收藏
Hugging Face2026-07-22 更新2026-07-22 收录
官方服务:

资源简介:

该数据集是一个用于机器人学研究的双手机器人操作数据集,由LeRobot工具创建,主要包含机器人执行折叠任务时的观察和动作数据。数据结构包括动作特征(16维浮点数组,表示左右机械臂的7个关节位置和夹爪位置)、观察状态(16维浮点数组,与动作特征相同),以及多个视频观察(基础视角:480x640分辨率,左右手腕视角:720x1280分辨率,均为30帧/秒的彩色视频)。此外,数据集还包含时间戳、帧索引、任务索引等元数据。数据集总共有2个任务,总帧数为4480帧,数据文件大小为100MB,视频文件大小为200MB。机器人类型为anvil_bimanual(双手机器人),适用于训练和评估机器人控制模型,特别是与折叠相关的任务。数据以Parquet格式存储,视频以MP4格式存储。

This dataset is a bimanual robot manipulation dataset for robotics research, created by the LeRobot tool. It primarily includes observation and action data from robots performing folding tasks. The data structure consists of action features (a 16-dimensional floating-point array representing the joint positions of the left and right arms with 7 joints each and gripper positions), observation states (a 16-dimensional floating-point array, same as the action features), and multiple video observations (base view: 480x640 resolution, left and right wrist views: 720x1280 resolution, all in color video at 30 frames per second). Additionally, the dataset contains metadata such as timestamps, frame indices, task indices, and more. There are a total of 2 tasks in the dataset, with 4480 frames in total, a data file size of 100MB, and a video file size of 200MB. The robot type is anvil_bimanual (bimanual robot), suitable for training and evaluating robot control models, especially for folding-related tasks. The data is stored in Parquet format, and videos are stored in MP4 format.

提供机构:
manual-cognition
创建时间:
2026-07-22
原始信息汇总

数据集概述

  • 数据集名称: rollout_folding_smolvla_full_20260722_150412
  • 许可证: Apache-2.0
  • 任务类别: 机器人学 (robotics)
  • 创建工具: 使用 LeRobot 创建

数据集结构

  • 数据格式: Parquet 和 MP4 视频文件
  • 帧率 (fps): 30
  • 总片段数: 2
  • 总帧数: 4480
  • 总任务数: 1
  • 数据文件大小: 约 100 MB
  • 视频文件大小: 约 200 MB
  • 机器人类型: anvil_bimanual
  • 数据集划分:
    • 训练集: 包含所有片段 (片段索引 0 到 1)

数据特征

数据集包含以下特征字段:

特征名称 数据类型 形状 描述
action float32 (16,) 机器人动作指令,包括左右臂各 7 个关节位置和 1 个夹爪位置
observation.state float32 (16,) 机器人观测状态,与动作维度一致
observation.images.base video (480, 640, 3) 基础视角摄像头视频,分辨率 480x640,AV1 编码,30 fps
observation.images.left_wrist video (720, 1280, 3) 左腕视角摄像头视频,分辨率 720x1280,AV1 编码,30 fps
observation.images.right_wrist video (720, 1280, 3) 右腕视角摄像头视频,分辨率 720x1280,AV1 编码,30 fps
timestamp float32 (1,) 时间戳
frame_index int64 (1,) 帧索引
episode_index int64 (1,) 片段索引
index int64 (1,) 索引
task_index int64 (1,) 任务索引

动作与状态空间

  • 动作维度: 16 维,对应左右臂各 7 个关节位置以及左右夹爪位置。
  • 观测状态: 16 维,与动作空间的物理含义一致。
  • 摄像头图像: 包含三个视角(基础、左腕、右腕),均为彩色视频。
二维码
社区交流群
二维码
科研交流群
商业服务