遇见数据集

Text prompts and videos generated using 5 popular Text-to-Video models plus quality metrics including user quality assessments

收藏
Mendeley Data2024-06-29 更新2024-06-28 收录
官方服务:

资源简介:

A collection of 201 prompts which are used to generate short-form videos using 5 popular text-to-video models namely Tune-a-Video, VideoFusion, Text-To-Vudeo Synthesis, Text2Video-Zero and Aphantasia. Each of the 1,005 generated videos is included along with automatically calculated quality metrics naturalness, text similarity between the original prompt and a generated text caption, and inception score, for each. Each video was rated by 24 different people and the data also includes the MOS scores for alignment between the generated videos and the original prompts, as well as for perception and overall quality of the video.Please cite this paper if using this dataset.

本数据集收录201条提示词,用于通过5款主流文本转视频模型生成短视频,这5款模型分别为Tune-a-Video、VideoFusion、文本到视频合成(Text-To-Video Synthesis)、Text2Video-Zero以及Aphantasia。数据集共包含1005段生成视频,每段视频均附带自动计算得到的三项质量评价指标:自然度、原始提示词与生成文本字幕间的文本相似度,以及初始评分(Inception Score)。每段视频均由24名不同受试者完成主观评分,数据集还包含生成视频与原始提示词对齐度的平均主观得分(Mean Opinion Score,简称MOS),以及视频感知质量与整体质量的MOS评分。若使用本数据集,请引用该论文。

创建时间:
2023-09-12
二维码
社区交流群
二维码
科研交流群
商业服务