官方服务:
资源简介:
This file contains scores of 9816 participants
应用场景:
创建时间:
2023-08-10
相关数据集
English–Tigrinya evaluation dataset
本研究构建了一个高质量的英语-提格雷语评估数据集,涵盖四个领域,包括宗教、新闻、健康和教育。数据集由1200个宗教领域句子、1500个新闻领域句子、800个健康领域句子和500个教育领域句子组成,总共有4000个句子。数据集已进行人工对齐和预处理,以支持严格的评估。
arXiv2025-09-24 更新330
shuyuej/French-MMLU-Anatomy-Benchmark
--- license: apache-2.0 --- # 💻 Dataset Usage Run the following command to load the testing set: ```python from datasets import load_dataset dataset = load_dataset("shuyuej/French-MMLU-Anatomy-Benc
Hugging Face2024-06-08 更新140
Mean identification accuracy (chance = 25%) by accent combination and focus location.
Accent combinations are in descending order of mean identification accuracy. Cell colouring signals the percentage of correct responses (25% < green < 50% < yellow < 75% < red) (colour online).
NIAID Data Ecosystem80
JFLEG: English Grammatical Error Benchmark
English Grammatical Error Correction Dataset
kaggle2023-12-02 更新160
Formal language assessment in low-educated healthy subjects
Abstract Although many studies have shown the influence of education on cognition, the impact of low education on the various cognitive functions appears to differ. The hypothesis of the present study
DataCite Commons2021-03-26 更新120



