遇见数据集

NJU-LINK/IF-VidCap

收藏
Hugging Face2025-10-22 更新2025-10-25 收录
官方服务:

资源简介:

IF-VidCap是一个用于评估视频字幕模型遵循指令能力的新基准,包含1400个高质量样本。它通过一个系统的框架评估字幕的格式正确性和内容正确性。数据集涵盖多种视频类别,并提供了一个健壮的评估协议以及详细的分析。

IF-VidCap is a new benchmark for evaluating the instruction-following capability of video captioning models, containing 1,400 high-quality samples. It assesses captions on two dimensions: format correctness and content correctness, with a robust evaluation protocol and detailed analysis across various video categories.

提供机构:
NJU-LINK
二维码
社区交流群
二维码
科研交流群
商业服务