遇见数据集

Multimodal-Fatima/VQAv2_test_3

收藏
Hugging Face2023-05-08 更新2024-03-04 收录
官方服务:

资源简介:

--- dataset_info: features: - name: question_type dtype: string - name: multiple_choice_answer dtype: string - name: answers sequence: string - name: answers_original list: - name: answer dtype: string - name: answer_confidence dtype: string - name: answer_id dtype: int64 - name: id_image dtype: int64 - name: answer_type dtype: string - name: question_id dtype: int64 - name: question dtype: string - name: image dtype: image - name: id dtype: int64 - name: clip_tags_ViT_L_14 sequence: string - name: blip_caption dtype: string - name: LLM_Description_gpt3_downstream_tasks_visual_genome_ViT_L_14 sequence: string - name: DETA_detections_deta_swin_large_o365_coco_classes list: - name: attribute dtype: string - name: box sequence: float32 - name: label dtype: string - name: location dtype: string - name: ratio dtype: float32 - name: size dtype: string - name: tag dtype: string - name: Attributes_ViT_L_14_descriptors_text_davinci_003_full sequence: string - name: clip_tags_ViT_L_14_wo_openai sequence: string - name: clip_tags_ViT_L_14_with_openai sequence: string - name: clip_tags_LAION_ViT_H_14_2B_wo_openai sequence: string - name: clip_tags_LAION_ViT_H_14_2B_with_openai sequence: string - name: clip_tags_LAION_ViT_bigG_14_2B_wo_openai sequence: string - name: clip_tags_LAION_ViT_bigG_14_2B_with_openai sequence: string - name: Attributes_LAION_ViT_H_14_2B_descriptors_text_davinci_003_full sequence: string - name: Attributes_LAION_ViT_bigG_14_2B_descriptors_text_davinci_003_full sequence: string splits: - name: test num_bytes: 14231060869.0 num_examples: 89558 download_size: 2658978296 dataset_size: 14231060869.0 --- # Dataset Card for "VQAv2_test_3" [More Information needed](https://github.com/huggingface/datasets/blob/main/CONTRIBUTING.md#how-to-contribute-to-the-dataset-cards)

数据集信息: 特征字段: - 名称:问题类型(question_type),数据类型:字符串 - 名称:单项选择答案(multiple_choice_answer),数据类型:字符串 - 名称:答案序列(answers),数据类型:字符串序列 - 名称:原始答案列表(answers_original),子字段: - 名称:答案(answer),数据类型:字符串 - 名称:答案置信度(answer_confidence),数据类型:字符串 - 名称:答案ID(answer_id),数据类型:int64 - 名称:图像ID(id_image),数据类型:int64 - 名称:答案类型(answer_type),数据类型:字符串 - 名称:问题ID(question_id),数据类型:int64 - 名称:问题文本(question),数据类型:字符串 - 名称:图像(image),数据类型:图像 - 名称:数据ID(id),数据类型:int64 - 名称:ViT-L/14的CLIP(Contrastive Language-Image Pre-training)标签序列(clip_tags_ViT_L_14),数据类型:字符串序列 - 名称:BLIP(Bootstrapping Language-Image Pre-training)图像描述(blip_caption),数据类型:字符串 - 名称:基于ViT-L/14与视觉基因组(Visual Genome)下游任务的GPT-3大语言模型(LLM)描述序列(LLM_Description_gpt3_downstream_tasks_visual_genome_ViT_L_14),数据类型:字符串序列 - 名称:基于DETA-Swin-Large模型、O365与COCO类别的DETA检测结果(DETA_detections_deta_swin_large_o365_coco_classes),子字段: - 名称:属性(attribute),数据类型:字符串 - 名称:边界框(box),数据类型:float32序列 - 名称:标签(label),数据类型:字符串 - 名称:位置(location),数据类型:字符串 - 名称:比例(ratio),数据类型:float32 - 名称:尺寸(size),数据类型:字符串 - 名称:标记(tag),数据类型:字符串 - 名称:使用text-davinci-003生成的ViT-L/14属性描述符序列(Attributes_ViT_L_14_descriptors_text_davinci_003_full),数据类型:字符串序列 - 名称:ViT-L/14的CLIP标签(不含OpenAI官方标签)序列(clip_tags_ViT_L_14_wo_openai),数据类型:字符串序列 - 名称:ViT-L/14的CLIP标签(含OpenAI官方标签)序列(clip_tags_ViT_L_14_with_openai),数据类型:字符串序列 - 名称:LAION-ViT-H/14 2B模型的CLIP标签(不含OpenAI官方标签)序列(clip_tags_LAION_ViT_H_14_2B_wo_openai),数据类型:字符串序列 - 名称:LAION-ViT-H/14 2B模型的CLIP标签(含OpenAI官方标签)序列(clip_tags_LAION_ViT_H_14_2B_with_openai),数据类型:字符串序列 - 名称:LAION-ViT-bigG/14 2B模型的CLIP标签(不含OpenAI官方标签)序列(clip_tags_LAION_ViT_bigG_14_2B_wo_openai),数据类型:字符串序列 - 名称:LAION-ViT-bigG/14 2B模型的CLIP标签(含OpenAI官方标签)序列(clip_tags_LAION_ViT_bigG_14_2B_with_openai),数据类型:字符串序列 - 名称:使用text-davinci-003生成的LAION-ViT-H/14 2B属性描述符序列(Attributes_LAION_ViT_H_14_2B_descriptors_text_davinci_003_full),数据类型:字符串序列 - 名称:使用text-davinci-003生成的LAION-ViT-bigG/14 2B属性描述符序列(Attributes_LAION_ViT_bigG_14_2B_descriptors_text_davinci_003_full),数据类型:字符串序列 数据集划分: - 名称:测试集(test),字节数:14231060869.0,样本数量:89558 下载大小:2658978296字节,数据集总大小:14231060869.0字节 # “VQAv2_test_3”数据集卡片 [需补充更多信息](https://github.com/huggingface/datasets/blob/main/CONTRIBUTING.md#how-to-contribute-to-the-dataset-cards)

提供机构:
Multimodal-Fatima
原始信息汇总

数据集概述

数据集名称

VQAv2_test_3

数据集特征

  • question_type: 字符串类型
  • multiple_choice_answer: 字符串类型
  • answers: 字符串序列
  • answers_original: 列表类型,包含:
    • answer: 字符串类型
    • answer_confidence: 字符串类型
    • answer_id: 整数类型
  • id_image: 整数类型
  • answer_type: 字符串类型
  • question_id: 整数类型
  • question: 字符串类型
  • image: 图像类型
  • id: 整数类型
  • clip_tags_ViT_L_14: 字符串序列
  • blip_caption: 字符串类型
  • LLM_Description_gpt3_downstream_tasks_visual_genome_ViT_L_14: 字符串序列
  • DETA_detections_deta_swin_large_o365_coco_classes: 列表类型,包含:
    • attribute: 字符串类型
    • box: 浮点数序列
    • label: 字符串类型
    • location: 字符串类型
    • ratio: 浮点数类型
    • size: 字符串类型
    • tag: 字符串类型
  • Attributes_ViT_L_14_descriptors_text_davinci_003_full: 字符串序列
  • clip_tags_ViT_L_14_wo_openai: 字符串序列
  • clip_tags_ViT_L_14_with_openai: 字符串序列
  • clip_tags_LAION_ViT_H_14_2B_wo_openai: 字符串序列
  • clip_tags_LAION_ViT_H_14_2B_with_openai: 字符串序列
  • clip_tags_LAION_ViT_bigG_14_2B_wo_openai: 字符串序列
  • clip_tags_LAION_ViT_bigG_14_2B_with_openai: 字符串序列
  • Attributes_LAION_ViT_H_14_2B_descriptors_text_davinci_003_full: 字符串序列
  • Attributes_LAION_ViT_bigG_14_2B_descriptors_text_davinci_003_full: 字符串序列

数据集分割

  • test: 包含89558个样本,数据集大小为14231060869.0字节。

数据集大小

  • 下载大小: 2658978296字节
  • 数据集总大小: 14231060869.0字节
搜集汇总
数据集介绍
Multimodal-Fatima/VQAv2_test_3 数据集图片
以上内容由遇见数据集搜集并总结生成
二维码
社区交流群
二维码
科研交流群
商业服务