human_v1_70f
收藏资源简介:
human_v1_70f是一个经过筛选和处理的多模态人类行为数据集,包含85,856个片段和343,424个上传文件。数据集包含四种文件类型:caption.pickle(字幕)、episode_meta.npz(元数据)、rgb.mp4(RGB视频)和skeleton_scenario.mp4(骨架场景视频)。数据来源于四个主要来源:egodex(78,273个片段)、vitra_ego4d_cooking(16个片段)、vitra_ego4d_other(13个片段)和vitra_epic(7,554个片段)。数据集在筛选时应用了以下标准:最小帧数(70帧)、最小可见比例(0.3)和最小手部运动量(0.005)。文件保留了原始处理数据的相对路径结构,并提供了原始分割清单和每个源数据集的保留片段ID列表。该数据集适用于人类行为分析、动作识别和多模态学习等任务。
human_v1_70f is a filtered and processed multimodal human behavior dataset containing 85,856 clips and 343,424 uploaded files. The dataset includes four types of files: caption.pickle (subtitles), episode_meta.npz (metadata), rgb.mp4 (RGB video), and skeleton_scenario.mp4 (skeleton scene video). The data comes from four main sources: egodex (78,273 clips), vitra_ego4d_cooking (16 clips), vitra_ego4d_other (13 clips), and vitra_epic (7,554 clips). The dataset was filtered using the following criteria: minimum number of frames (70 frames), minimum visible ratio (0.3), and minimum hand movement (0.005). The files retain the relative path structure of the original processed data and provide the original split manifest and a list of retained clip IDs for each source dataset. This dataset is suitable for tasks such as human behavior analysis, action recognition, and multimodal learning.
数据集概述:masterwu/human_v1_70f
基本信息
- 数据集名称:masterwu/human_v1_70f
- 许可证:其他(license: other)
- 数据规模:包含 85,856 个 episodes
文件内容
数据集包含以下4种文件类型,每种文件数量均为 85,856 个:
rgb.mp4:RGB 视频文件episode_meta.npz:片段元数据caption.pickle:描述文本skeleton_scenario.mp4:骨骼场景视频
已上传文件总数:343,424 个
数据来源
数据集由多个来源组成,具体如下:
- egodex:78,273 episodes
- vitra_epic:7,554 episodes
- vitra_ego4d_cooking:16 episodes
- vitra_ego4d_other:13 episodes
过滤条件
对原始数据应用了以下过滤条件:
min_frames:70(最小帧数)min_visible_ratio:0.3(最小可见比例)min_hand_motion:0.005(最小手部运动量)
文件布局
- 片段文件保留在原始处理后的数据根目录下的相对路径
human_train_manifest_v1.json:原始分割清单文件manifests/*.manifest.json:按来源数据集列出保留的片段 IDrsync_files.txt:权威文件列表,用于upload_mode=all
文件生成时间
数据集文件生成于:2026-04-22 12:33:49




