遇见数据集

ImageNet-VidVRD

收藏
arXiv2025-09-30 收录
官方服务:

资源简介:

该数据集是从ILSVRC2016-VID中筛选出的一部分,包含了展现清晰视觉关系的视频。它涵盖了35个物体类别和132个谓词类别,并带有片段级别的视觉关系标注。规模上,该数据集包含了1000个视频,其中800个用于训练,200个用于测试。其任务是视频视觉关系检测(Vidvrd)。

This dataset is a curated subset selected from ILSVRC2016-VID, comprising videos that demonstrate clear visual relationships. It includes 35 object categories and 132 predicate categories, with clip-level visual relationship annotations. In terms of scale, the dataset consists of 1000 videos in total, where 800 videos are allocated for training and the remaining 200 for testing. The corresponding task is Video Visual Relationship Detection (Vidvrd).

提供机构:
ILSVRC2016
搜集汇总
数据集介绍
ImageNet-VidVRD 数据集图片
背景与挑战
背景概述
ImageNet-VidVRD是首个视频视觉关系检测(VidVRD)数据集,包含1000个从ILVSRC2016-VID选出的视频,分为800个训练和200个测试视频,覆盖35个对象类别和132个关系谓词类别。该数据集旨在通过标注对象轨迹和视觉关系实例(如<subject, predicate, object>三元组),促进视频中动态和时序变化关系的研究,例如训练集有2,961个关系三元组,测试集有1,011个三元组和4,835个实例。
以上内容由遇见数据集搜集并总结生成
二维码
社区交流群
二维码
科研交流群
商业服务