CV-Bench
收藏资源简介:
该数据集是一个多模态数据集,包含2638个测试样本。每个样本由三个主要字段构成:唯一标识符(id字段,字符串类型)、图像媒体内容(media字段,图像列表形式)以及文本消息(messages字段,字符串类型)。数据集总大小约为346MB,仅提供测试集分割。从数据结构推断,该数据集适用于需要结合图像和文本信息的任务,例如视觉问答、图文对话或多模态理解等应用场景。
This dataset is a multimodal dataset containing 2638 test samples. Each sample consists of three main fields: a unique identifier (id field, string type), image media content (media field, in the form of an image list), and text messages (messages field, string type). The total dataset size is approximately 346MB, and only a test set split is provided. Based on the data structure, this dataset is suitable for tasks that require combining image and text information, such as visual question answering, image-text dialogue, or multimodal understanding applications.
数据集概览
- 数据集名称: CV-Bench
- 数据集页面: https://huggingface.co/datasets/mm-eval/CV-Bench
数据特征
该数据集包含以下特征字段:
- id: 字符串类型,表示样本的唯一标识符。
- media: 图像列表,包含多个图像。
- messages: 字符串类型,可能包含与任务相关的指令或对话内容。
数据划分
- 数据集仅包含一个划分:test(测试集)。
- 测试集包含 2638 个样本,总字节数为 345,969,926(约 330 MB)。
数据集大小
- 下载大小: 345,180,584 字节(约 329 MB)
- 数据集总大小: 345,969,926 字节(约 330 MB)




