Multimodal-Fatima/OK-VQA_train_embeddings
收藏官方服务:
资源简介:
--- dataset_info: features: - name: image dtype: image - name: id dtype: int64 - name: vision_embeddings sequence: float32 splits: - name: openai_clip_vit_large_patch14 num_bytes: 1513678502.0 num_examples: 9009 download_size: 1517323156 dataset_size: 1513678502.0 --- # Dataset Card for "OK-VQA_train_embeddings" [More Information needed](https://github.com/huggingface/datasets/blob/main/CONTRIBUTING.md#how-to-contribute-to-the-dataset-cards)
提供机构:
Multimodal-Fatima原始信息汇总
数据集概述
数据集名称
OK-VQA_train_embeddings
数据集特征
- image: 图像数据
- id: 整数类型,可能用于标识图像
- vision_embeddings: 浮点数序列,可能包含图像的视觉嵌入信息
数据集分割
- openai_clip_vit_large_patch14:
- 数据量: 1513678502.0 字节
- 样本数: 9009
数据集大小
- 下载大小: 1517323156 字节
- 实际数据集大小: 1513678502.0 字节



