Multimodal-Fatima/VQAv2_test_3
收藏资源简介:
--- dataset_info: features: - name: question_type dtype: string - name: multiple_choice_answer dtype: string - name: answers sequence: string - name: answers_original list: - name: answer dtype: string - name: answer_confidence dtype: string - name: answer_id dtype: int64 - name: id_image dtype: int64 - name: answer_type dtype: string - name: question_id dtype: int64 - name: question dtype: string - name: image dtype: image - name: id dtype: int64 - name: clip_tags_ViT_L_14 sequence: string - name: blip_caption dtype: string - name: LLM_Description_gpt3_downstream_tasks_visual_genome_ViT_L_14 sequence: string - name: DETA_detections_deta_swin_large_o365_coco_classes list: - name: attribute dtype: string - name: box sequence: float32 - name: label dtype: string - name: location dtype: string - name: ratio dtype: float32 - name: size dtype: string - name: tag dtype: string - name: Attributes_ViT_L_14_descriptors_text_davinci_003_full sequence: string - name: clip_tags_ViT_L_14_wo_openai sequence: string - name: clip_tags_ViT_L_14_with_openai sequence: string - name: clip_tags_LAION_ViT_H_14_2B_wo_openai sequence: string - name: clip_tags_LAION_ViT_H_14_2B_with_openai sequence: string - name: clip_tags_LAION_ViT_bigG_14_2B_wo_openai sequence: string - name: clip_tags_LAION_ViT_bigG_14_2B_with_openai sequence: string - name: Attributes_LAION_ViT_H_14_2B_descriptors_text_davinci_003_full sequence: string - name: Attributes_LAION_ViT_bigG_14_2B_descriptors_text_davinci_003_full sequence: string splits: - name: test num_bytes: 14231060869.0 num_examples: 89558 download_size: 2658978296 dataset_size: 14231060869.0 --- # Dataset Card for "VQAv2_test_3" [More Information needed](https://github.com/huggingface/datasets/blob/main/CONTRIBUTING.md#how-to-contribute-to-the-dataset-cards)
数据集信息: 特征字段: - 名称:问题类型(question_type),数据类型:字符串 - 名称:单项选择答案(multiple_choice_answer),数据类型:字符串 - 名称:答案序列(answers),数据类型:字符串序列 - 名称:原始答案列表(answers_original),子字段: - 名称:答案(answer),数据类型:字符串 - 名称:答案置信度(answer_confidence),数据类型:字符串 - 名称:答案ID(answer_id),数据类型:int64 - 名称:图像ID(id_image),数据类型:int64 - 名称:答案类型(answer_type),数据类型:字符串 - 名称:问题ID(question_id),数据类型:int64 - 名称:问题文本(question),数据类型:字符串 - 名称:图像(image),数据类型:图像 - 名称:数据ID(id),数据类型:int64 - 名称:ViT-L/14的CLIP(Contrastive Language-Image Pre-training)标签序列(clip_tags_ViT_L_14),数据类型:字符串序列 - 名称:BLIP(Bootstrapping Language-Image Pre-training)图像描述(blip_caption),数据类型:字符串 - 名称:基于ViT-L/14与视觉基因组(Visual Genome)下游任务的GPT-3大语言模型(LLM)描述序列(LLM_Description_gpt3_downstream_tasks_visual_genome_ViT_L_14),数据类型:字符串序列 - 名称:基于DETA-Swin-Large模型、O365与COCO类别的DETA检测结果(DETA_detections_deta_swin_large_o365_coco_classes),子字段: - 名称:属性(attribute),数据类型:字符串 - 名称:边界框(box),数据类型:float32序列 - 名称:标签(label),数据类型:字符串 - 名称:位置(location),数据类型:字符串 - 名称:比例(ratio),数据类型:float32 - 名称:尺寸(size),数据类型:字符串 - 名称:标记(tag),数据类型:字符串 - 名称:使用text-davinci-003生成的ViT-L/14属性描述符序列(Attributes_ViT_L_14_descriptors_text_davinci_003_full),数据类型:字符串序列 - 名称:ViT-L/14的CLIP标签(不含OpenAI官方标签)序列(clip_tags_ViT_L_14_wo_openai),数据类型:字符串序列 - 名称:ViT-L/14的CLIP标签(含OpenAI官方标签)序列(clip_tags_ViT_L_14_with_openai),数据类型:字符串序列 - 名称:LAION-ViT-H/14 2B模型的CLIP标签(不含OpenAI官方标签)序列(clip_tags_LAION_ViT_H_14_2B_wo_openai),数据类型:字符串序列 - 名称:LAION-ViT-H/14 2B模型的CLIP标签(含OpenAI官方标签)序列(clip_tags_LAION_ViT_H_14_2B_with_openai),数据类型:字符串序列 - 名称:LAION-ViT-bigG/14 2B模型的CLIP标签(不含OpenAI官方标签)序列(clip_tags_LAION_ViT_bigG_14_2B_wo_openai),数据类型:字符串序列 - 名称:LAION-ViT-bigG/14 2B模型的CLIP标签(含OpenAI官方标签)序列(clip_tags_LAION_ViT_bigG_14_2B_with_openai),数据类型:字符串序列 - 名称:使用text-davinci-003生成的LAION-ViT-H/14 2B属性描述符序列(Attributes_LAION_ViT_H_14_2B_descriptors_text_davinci_003_full),数据类型:字符串序列 - 名称:使用text-davinci-003生成的LAION-ViT-bigG/14 2B属性描述符序列(Attributes_LAION_ViT_bigG_14_2B_descriptors_text_davinci_003_full),数据类型:字符串序列 数据集划分: - 名称:测试集(test),字节数:14231060869.0,样本数量:89558 下载大小:2658978296字节,数据集总大小:14231060869.0字节 # “VQAv2_test_3”数据集卡片 [需补充更多信息](https://github.com/huggingface/datasets/blob/main/CONTRIBUTING.md#how-to-contribute-to-the-dataset-cards)
数据集概述
数据集名称
VQAv2_test_3
数据集特征
- question_type: 字符串类型
- multiple_choice_answer: 字符串类型
- answers: 字符串序列
- answers_original: 列表类型,包含:
- answer: 字符串类型
- answer_confidence: 字符串类型
- answer_id: 整数类型
- id_image: 整数类型
- answer_type: 字符串类型
- question_id: 整数类型
- question: 字符串类型
- image: 图像类型
- id: 整数类型
- clip_tags_ViT_L_14: 字符串序列
- blip_caption: 字符串类型
- LLM_Description_gpt3_downstream_tasks_visual_genome_ViT_L_14: 字符串序列
- DETA_detections_deta_swin_large_o365_coco_classes: 列表类型,包含:
- attribute: 字符串类型
- box: 浮点数序列
- label: 字符串类型
- location: 字符串类型
- ratio: 浮点数类型
- size: 字符串类型
- tag: 字符串类型
- Attributes_ViT_L_14_descriptors_text_davinci_003_full: 字符串序列
- clip_tags_ViT_L_14_wo_openai: 字符串序列
- clip_tags_ViT_L_14_with_openai: 字符串序列
- clip_tags_LAION_ViT_H_14_2B_wo_openai: 字符串序列
- clip_tags_LAION_ViT_H_14_2B_with_openai: 字符串序列
- clip_tags_LAION_ViT_bigG_14_2B_wo_openai: 字符串序列
- clip_tags_LAION_ViT_bigG_14_2B_with_openai: 字符串序列
- Attributes_LAION_ViT_H_14_2B_descriptors_text_davinci_003_full: 字符串序列
- Attributes_LAION_ViT_bigG_14_2B_descriptors_text_davinci_003_full: 字符串序列
数据集分割
- test: 包含89558个样本,数据集大小为14231060869.0字节。
数据集大小
- 下载大小: 2658978296字节
- 数据集总大小: 14231060869.0字节




