遇见数据集

rh20t_cfg7

收藏
Hugging Face2026-06-21 更新2026-06-21 收录
官方服务:

资源简介:

rh20t_cfg7是一个基于RH20T配置的非官方LeRobot数据集v3格式转换版本,专注于机器人操作任务。它包含896个episodes和385,158帧图像,仅提供RGB图像(无深度信息),帧率为10fps,使用Kuka类型机器人,并包含10个相机流。数据来源于RH20T数据集(Fang等人,arXiv:2307.00595),保留了原始的双重许可:场景0001-0005使用CC BY-SA 4.0许可(允许商业使用,需署名和以相同方式共享),场景0006-0010使用CC BY-NC 4.0许可(仅限非商业使用)。数据集包含志愿者记录的人类-机器人交互内容,可能涉及人脸(视频)和语音(音频),需谨慎使用。数据模式包括观察状态(如末端执行器姿态、关节位置、夹持器状态)、力和扭矩数据、动作(如下一末端执行器姿态和夹持器状态)、元数据评分以及10个相机流的视频图像。

rh20t_cfg7 is an unofficial LeRobot dataset v3 format conversion version based on the RH20T configuration, focusing on robotic manipulation tasks. It contains 896 episodes and 385,158 image frames, providing only RGB images (no depth), with a frame rate of 10fps, using a Kuka-type robot, and includes 10 camera streams. The data originates from the RH20T dataset (Fang et al., arXiv:2307.00595) and retains the original dual license: scenes 0001-0005 use the CC BY-SA 4.0 license (allowing commercial use with attribution and share-alike), while scenes 0006-0010 use the CC BY-NC 4.0 license (non-commercial use only). The dataset includes human-robot interaction content recorded by volunteers, which may involve faces (video) and speech (audio), requiring careful use. Data modalities include observation states (such as end-effector pose, joint positions, gripper state), force and torque data, actions (such as next end-effector pose and gripper state), metadata scores, and video images from 10 camera streams.

提供机构:
robot-lev
创建时间:
2026-06-21
原始信息汇总

数据集概述

rh20t_cfg7 是 RH20T 数据集的一个非官方 LeRobot Dataset v3 重格式化版本,专注于机器人操作任务。

基本信息

  • 任务类别: 机器人学 (robotics)
  • 数据集格式: LeRobot v3
  • 内容类型: 仅 RGB 图像(无深度信息)
  • 机器人类型: kuka

数据规模

  • 片段数 (episodes): 896
  • 总帧数: 385,158
  • 帧率: 10 fps
  • 摄像头流: 10 个摄像头流

数据构成

每个数据片段包含以下字段:

  • observation.state: 末端执行器位姿 (7维) + 关节位置 (7维) + 夹爪状态 (1维)
  • observation.forceobservation.torque: 力和力矩 (各3维)
  • observation.robot_ft: 机器人末端力/扭矩 (6维)
  • action: 下一时刻的末端执行器位姿 (7维) + 夹爪状态 (1维)
  • meta.rating: 评分信息
  • observation.images.cam_<serial>: 10 个摄像头流的视频图像

许可协议

数据集采用基于场景的双重许可

  • scene_0001 至 scene_0005: CC BY-SA 4.0(允许商业使用,需署名并相同方式共享)
  • scene_0006 至 scene_0010: CC BY-NC 4.0(仅非商业使用,禁止商业用途)
  • 每个片段的具体许可可通过 meta/rh20t_episodes.json 中的 folder 字段确定

敏感内容提示

数据集包含志愿者录制的人机交互内容,可能包含面部(视频)和声音(音频副文件)。请谨慎处理,避免查看或分享敏感内容,仅用于模型训练目的。

引用来源

原始论文: Fang et al., RH20T: A Comprehensive Robotic Dataset for Learning Diverse Skills in One-Shot, arXiv:2307.00595

相关链接

  • 项目主页: https://rh20t.github.io/
  • RH20T API: https://github.com/rh20t/rh20t_api
  • 重格式化代码: https://github.com/lvjonok/rh20t-lerobot-port
二维码
社区交流群
二维码
科研交流群
商业服务