MSRVTT-I2V
收藏资源简介:
MSRVTT-I2V是一个基于MSR-VTT视频数据集构建的图像到视频检索任务专用数据集,来源于UVRB基准,旨在为跨模态检索研究提供标准化的评估数据。数据内容包含查询集(由1,000个从原始视频中预留出的独立视频帧构成)和语料库(对应的1,000个完整的MSR-VTT视频),采用MTEB标准检索格式进行组织,包括语料库文档、查询和相关性判断等标准组件,便于在统一框架下进行检索模型的训练与评估。该数据集适用于图像-视频跨模态检索、表示学习等相关研究任务。
MSRVTT-I2V is a dedicated image-to-video retrieval dataset constructed based on the MSR-VTT video dataset, derived from the UVRB benchmark, aiming to provide standardized evaluation data for cross-modal retrieval research. The dataset includes a query set (composed of 1,000 independent video frames reserved from the original videos) and a corpus (the corresponding 1,000 complete MSR-VTT videos). It is organized in the MTEB standard retrieval format, covering standard components such as corpus documents, queries and relevance judgments, which facilitates the training and evaluation of retrieval models under a unified framework. This dataset is suitable for research tasks such as image-video cross-modal retrieval and representation learning.
数据集概述:MSRVTT-I2V
该数据集是基于 MSR-VTT 构建的图像到视频检索分割数据集,源自 UVRB 基准(arXiv:2510.27571)。
- 用途:图像到视频检索(Image-to-video retrieval)。
- 规模:
- 查询(Queries):1,000 个保留的视频帧。
- 语料库(Corpus):对应的 1,000 个 MSR-VTT 视频。
- 格式:采用 MTEB 标准检索格式,包含四个配置:
- corpus:语料库,测试集(test)。
- queries:查询,测试集(test)。
- pc:测试集(test)。
- qrels:相关度判断,测试集(test)。
- 许可证:其他(msr-vtt-research),详情请参见微软研究院(https://www.microsoft.com/en-us/research/)。





