gagan3012/multilingual-llava-bench-in-the-wild
收藏资源简介:
--- dataset_info: - config_name: ar features: - name: question_id dtype: int64 - name: image dtype: image - name: question dtype: string - name: caption dtype: string - name: image_id dtype: string - name: gpt_answer dtype: string - name: category dtype: string splits: - name: test num_bytes: 22342774.0 num_examples: 60 download_size: 9778993 dataset_size: 22342774.0 - config_name: arabic features: - name: question_id dtype: int64 - name: image dtype: image - name: question dtype: string - name: caption dtype: string - name: image_id dtype: string - name: gpt_answer dtype: string - name: category dtype: string splits: - name: test num_bytes: 22342774.0 num_examples: 60 download_size: 9778993 dataset_size: 22342774.0 - config_name: bengali features: - name: question_id dtype: int64 - name: image dtype: image - name: question dtype: string - name: caption dtype: string - name: image_id dtype: string - name: gpt_answer dtype: string - name: category dtype: string splits: - name: test num_bytes: 22378020.0 num_examples: 60 download_size: 9783130 dataset_size: 22378020.0 - config_name: chinese features: - name: question_id dtype: int64 - name: image dtype: image - name: question dtype: string - name: caption dtype: string - name: image_id dtype: string - name: gpt_answer dtype: string - name: category dtype: string splits: - name: test num_bytes: 22317502.0 num_examples: 60 download_size: 9772605 dataset_size: 22317502.0 - config_name: french features: - name: question_id dtype: int64 - name: image dtype: image - name: question dtype: string - name: caption dtype: string - name: image_id dtype: string - name: gpt_answer dtype: string - name: category dtype: string splits: - name: test num_bytes: 22327391.0 num_examples: 60 download_size: 9773783 dataset_size: 22327391.0 - config_name: hindi features: - name: question_id dtype: int64 - name: image dtype: image - name: question dtype: string - name: caption dtype: string - name: image_id dtype: string - name: gpt_answer dtype: string - name: category dtype: string splits: - name: test num_bytes: 22385129.0 num_examples: 60 download_size: 9799590 dataset_size: 22385129.0 - config_name: japanese features: - name: question_id dtype: int64 - name: image dtype: image - name: question dtype: string - name: caption dtype: string - name: image_id dtype: string - name: gpt_answer dtype: string - name: category dtype: string splits: - name: test num_bytes: 22333016.0 num_examples: 60 download_size: 9782382 dataset_size: 22333016.0 - config_name: russian features: - name: question_id dtype: int64 - name: image dtype: image - name: question dtype: string - name: caption dtype: string - name: image_id dtype: string - name: gpt_answer dtype: string - name: category dtype: string splits: - name: test num_bytes: 22355236.0 num_examples: 60 download_size: 9792575 dataset_size: 22355236.0 - config_name: spanish features: - name: question_id dtype: int64 - name: image dtype: image - name: question dtype: string - name: caption dtype: string - name: image_id dtype: string - name: gpt_answer dtype: string - name: category dtype: string splits: - name: test num_bytes: 22326471.0 num_examples: 60 download_size: 9781970 dataset_size: 22326471.0 - config_name: urdu features: - name: question_id dtype: int64 - name: image dtype: image - name: question dtype: string - name: caption dtype: string - name: image_id dtype: string - name: gpt_answer dtype: string - name: category dtype: string splits: - name: test num_bytes: 22349409.0 num_examples: 60 download_size: 9784751 dataset_size: 22349409.0 configs: - config_name: ar data_files: - split: test path: ar/test-* - config_name: arabic data_files: - split: test path: arabic/test-* - config_name: bengali data_files: - split: test path: bengali/test-* - config_name: chinese data_files: - split: test path: chinese/test-* - config_name: french data_files: - split: test path: french/test-* - config_name: hindi data_files: - split: test path: hindi/test-* - config_name: japanese data_files: - split: test path: japanese/test-* - config_name: russian data_files: - split: test path: russian/test-* - config_name: spanish data_files: - split: test path: spanish/test-* - config_name: urdu data_files: - split: test path: urdu/test-* ---
数据集信息: - 配置名称:ar(阿拉伯语) 特征: - 名称:问题ID,数据类型:64位整数 - 名称:图像,数据类型:图像 - 名称:问题,数据类型:字符串 - 名称:图像说明,数据类型:字符串 - 名称:图像ID,数据类型:字符串 - 名称:GPT生成回答,数据类型:字符串 - 名称:类别,数据类型:字符串 数据集划分: - 划分名称:测试集,字节数:22342774.0,样本数量:60 下载大小:9778993,数据集总大小:22342774.0 - 配置名称:arabic(阿拉伯语) 特征: - 名称:问题ID,数据类型:64位整数 - 名称:图像,数据类型:图像 - 名称:问题,数据类型:字符串 - 名称:图像说明,数据类型:字符串 - 名称:图像ID,数据类型:字符串 - 名称:GPT生成回答,数据类型:字符串 - 名称:类别,数据类型:字符串 数据集划分: - 划分名称:测试集,字节数:22342774.0,样本数量:60 下载大小:9778993,数据集总大小:22342774.0 - 配置名称:bengali(孟加拉语) 特征: - 名称:问题ID,数据类型:64位整数 - 名称:图像,数据类型:图像 - 名称:问题,数据类型:字符串 - 名称:图像说明,数据类型:字符串 - 名称:图像ID,数据类型:字符串 - 名称:GPT生成回答,数据类型:字符串 - 名称:类别,数据类型:字符串 数据集划分: - 划分名称:测试集,字节数:22378020.0,样本数量:60 下载大小:9783130,数据集总大小:22378020.0 - 配置名称:chinese(中文) 特征: - 名称:问题ID,数据类型:64位整数 - 名称:图像,数据类型:图像 - 名称:问题,数据类型:字符串 - 名称:图像说明,数据类型:字符串 - 名称:图像ID,数据类型:字符串 - 名称:GPT生成回答,数据类型:字符串 - 名称:类别,数据类型:字符串 数据集划分: - 划分名称:测试集,字节数:22317502.0,样本数量:60 下载大小:9772605,数据集总大小:22317502.0 - 配置名称:french(法语) 特征: - 名称:问题ID,数据类型:64位整数 - 名称:图像,数据类型:图像 - 名称:问题,数据类型:字符串 - 名称:图像说明,数据类型:字符串 - 名称:图像ID,数据类型:字符串 - 名称:GPT生成回答,数据类型:字符串 - 名称:类别,数据类型:字符串 数据集划分: - 划分名称:测试集,字节数:22327391.0,样本数量:60 下载大小:9773783,数据集总大小:22327391.0 - 配置名称:hindi(印地语) 特征: - 名称:问题ID,数据类型:64位整数 - 名称:图像,数据类型:图像 - 名称:问题,数据类型:字符串 - 名称:图像说明,数据类型:字符串 - 名称:图像ID,数据类型:字符串 - 名称:GPT生成回答,数据类型:字符串 - 名称:类别,数据类型:字符串 数据集划分: - 划分名称:测试集,字节数:22385129.0,样本数量:60 下载大小:9799590,数据集总大小:22385129.0 - 配置名称:japanese(日语) 特征: - 名称:问题ID,数据类型:64位整数 - 名称:图像,数据类型:图像 - 名称:问题,数据类型:字符串 - 名称:图像说明,数据类型:字符串 - 名称:图像ID,数据类型:字符串 - 名称:GPT生成回答,数据类型:字符串 - 名称:类别,数据类型:字符串 数据集划分: - 划分名称:测试集,字节数:22333016.0,样本数量:60 下载大小:9782382,数据集总大小:22333016.0 - 配置名称:russian(俄语) 特征: - 名称:问题ID,数据类型:64位整数 - 名称:图像,数据类型:图像 - 名称:问题,数据类型:字符串 - 名称:图像说明,数据类型:字符串 - 名称:图像ID,数据类型:字符串 - 名称:GPT生成回答,数据类型:字符串 - 名称:类别,数据类型:字符串 数据集划分: - 划分名称:测试集,字节数:22355236.0,样本数量:60 下载大小:9792575,数据集总大小:22355236.0 - 配置名称:spanish(西班牙语) 特征: - 名称:问题ID,数据类型:64位整数 - 名称:图像,数据类型:图像 - 名称:问题,数据类型:字符串 - 名称:图像说明,数据类型:字符串 - 名称:图像ID,数据类型:字符串 - 名称:GPT生成回答,数据类型:字符串 - 名称:类别,数据类型:字符串 数据集划分: - 划分名称:测试集,字节数:22326471.0,样本数量:60 下载大小:9781970,数据集总大小:22326471.0 - 配置名称:urdu(乌尔都语) 特征: - 名称:问题ID,数据类型:64位整数 - 名称:图像,数据类型:图像 - 名称:问题,数据类型:字符串 - 名称:图像说明,数据类型:字符串 - 名称:图像ID,数据类型:字符串 - 名称:GPT生成回答,数据类型:字符串 - 名称:类别,数据类型:字符串 数据集划分: - 划分名称:测试集,字节数:22349409.0,样本数量:60 下载大小:9784751,数据集总大小:22349409.0 配置列表: - 配置名称:ar 数据文件: - 划分:测试集,路径:ar/test-* - 配置名称:arabic 数据文件: - 划分:测试集,路径:arabic/test-* - 配置名称:bengali 数据文件: - 划分:测试集,路径:bengali/test-* - 配置名称:chinese 数据文件: - 划分:测试集,路径:chinese/test-* - 配置名称:french 数据文件: - 划分:测试集,路径:french/test-* - 配置名称:hindi 数据文件: - 划分:测试集,路径:hindi/test-* - 配置名称:japanese 数据文件: - 划分:测试集,路径:japanese/test-* - 配置名称:russian 数据文件: - 划分:测试集,路径:russian/test-* - 配置名称:spanish 数据文件: - 划分:测试集,路径:spanish/test-* - 配置名称:urdu 数据文件: - 划分:测试集,路径:urdu/test-*
数据集概述
数据集配置及特征
- 配置名称: ar, arabic, bengali, chinese, french, hindi, japanese, russian, spanish, urdu
- 特征:
- question_id: 数据类型为 int64
- image: 数据类型为 image
- question: 数据类型为 string
- caption: 数据类型为 string
- image_id: 数据类型为 string
- gpt_answer: 数据类型为 string
- category: 数据类型为 string
数据集分割
- 分割名称: test
- 示例数量: 60
- 字节数:
- ar, arabic: 22342774.0
- bengali: 22378020.0
- chinese: 22317502.0
- french: 22327391.0
- hindi: 22385129.0
- japanese: 22333016.0
- russian: 22355236.0
- spanish: 22326471.0
- urdu: 22349409.0
数据集大小及下载大小
- 下载大小:
- ar, arabic: 9778993
- bengali: 9783130
- chinese: 9772605
- french: 9773783
- hindi: 9799590
- japanese: 9782382
- russian: 9792575
- spanish: 9781970
- urdu: 9784751
- 数据集大小: 与字节数相同
数据文件路径
- 配置名称: ar, arabic, bengali, chinese, french, hindi, japanese, russian, spanish, urdu
- 分割: test
- 路径:
- ar: ar/test-*
- arabic: arabic/test-*
- bengali: bengali/test-*
- chinese: chinese/test-*
- french: french/test-*
- hindi: hindi/test-*
- japanese: japanese/test-*
- russian: russian/test-*
- spanish: spanish/test-*
- urdu: urdu/test-*




