davidadamczyk/czechbench_belebele
收藏资源简介:
--- dataset_info: features: - name: link dtype: string - name: question_number dtype: int64 - name: flores_passage dtype: string - name: question dtype: string - name: mc_answer1 dtype: string - name: mc_answer2 dtype: string - name: mc_answer3 dtype: string - name: mc_answer4 dtype: string - name: correct_answer_num dtype: string - name: dialect dtype: string - name: ds dtype: timestamp[us] splits: - name: train num_bytes: 15940.0 num_examples: 20 - name: test num_bytes: 735917.0 num_examples: 880 download_size: 364412 dataset_size: 751857.0 configs: - config_name: default data_files: - split: train path: data/train-* - split: test path: data/test-* --- # Dataset Card for "czechbench_belebele" [More Information needed](https://github.com/huggingface/datasets/blob/main/CONTRIBUTING.md#how-to-contribute-to-the-dataset-cards)
数据集概述
数据集名称
czechbench_belebele
数据集特征
- link: 字符串类型
- question_number: 整数类型
- flores_passage: 字符串类型
- question: 字符串类型
- mc_answer1: 字符串类型
- mc_answer2: 字符串类型
- mc_answer3: 字符串类型
- mc_answer4: 字符串类型
- correct_answer_num: 字符串类型
- dialect: 字符串类型
- ds: 时间戳类型,单位为微秒
数据集分割
- test:
- 数据量: 735917字节
- 示例数量: 880
- train:
- 数据量: 15940字节
- 示例数量: 20
数据集大小
- 下载大小: 363794字节
- 数据集大小: 751857字节



