VideoEval-Pro
收藏资源简介:
VideoEval-Pro是一个稳健且现实的长视频理解基准测试,包含开放式、短答案问答问题。该数据集通过将来自四个现有的长视频理解多项选择基准测试(Video-MME、MLVU、LVBench和LongVideoBench)的问题重新格式化为自由形式的问题来构建。
VideoEval-Pro represents a robust and realistic long video understanding benchmark test, encompassing open-ended and short-answer question and answer scenarios. The dataset is constructed by reformatting questions from four existing long video understanding multiple-choice benchmarks (Video-MME, MLVU, LVBench, and LongVideoBench) into free-form questions.
VideoEval-Pro 数据集概述
数据集简介
- 名称:VideoEval-Pro
- 类型:长视频理解评估基准
- 特点:包含开放式短答案问答问题,通过重构四个现有长视频理解多选题基准(Video-MME、MLVU、LVBench、LongVideoBench)的问题构建而成
数据内容
每个数据样本包含以下字段:
video:视频文件名(路径)question:关于视频内容的问题options:原始多选题选项answer:正确的多选题答案answer_text:正确的自由形式答案meta:来自源基准的额外元数据source:源基准名称qa_subtype:问题任务子类型qa_type:问题任务类型
数据获取
- 下载地址:https://huggingface.co/datasets/TIGER-Lab/VideoEval-Pro
- 下载方式: bash git lfs install git clone https://huggingface.co/datasets/TIGER-Lab/VideoEval-Pro
视频处理
-
合并视频分片: bash cat videos_part_*.tar.gz > videos_merged.tar.gz
-
解压视频: bash tar -xzf videos_merged.tar.gz
可选帧提取
-
目录结构:
frames_root/ ├── video_name_1/ │ ├── 000001.jpg │ └── ... └── ...
评估环境配置
bash git clone https://github.com/TIGER-AI-Lab/VideoEval-Pro cd VideoEval-Pro conda create -n videoevalpro --file *.yaml conda activate videoevalpro
评估运行
bash python tools/*_chat.py --video_root <path_to_videos> --frames_root <path_to_frames> --output_path <path_to_save_results> --using_frames <True/False> --model_path <model_name_or_path> --device <device> --num_frames <number_of_frames> --max_retries <max_retries> --num_threads <num_threads>




