LehongWu/vis-sft_rl_step680-val30-0529
收藏资源简介:
该数据集是一个多模态数据集,包含图像和文本数据,用于评估或训练模型在结合视觉和语言任务中的性能。数据集特征包括图像列表、提示(包含内容和角色信息)、奖励模型(包含真实标签和风格信息)以及额外信息(如答案、完成内容、思考过程、唯一标识符、目标、任务特定提示和先前指令)。数据来源、能力类型、数据划分和模型完成内容也作为字段提供。数据集仅包含测试分集,共有30个样本,适用于多模态人工智能任务的研究和评估。
This dataset is a multimodal dataset containing images and text data, designed for evaluating or training models on tasks that integrate vision and language. Features include a list of images, prompts (with content and role information), reward models (with ground truth and style information), and extra information (such as answer, completion, think process, unique identifier, goal, task-specific prompt, and previous instruction). Data source, ability type, data split, and model completion are also provided as fields. The dataset includes only a test split with 30 samples, suitable for research and evaluation in multimodal AI tasks.




