遇见数据集

A Multilingual Dataset of Student Answers, Human Grading, and Multi-LLM Evaluations for Automated Assessment Research using JorGPT

收藏
Zenodo2026-03-24 更新2026-05-26 收录
官方服务:

资源简介:

This is a multilingual dataset of open-ended student answers collected from authentic university-level assessments. The dataset comprises 3041 student responses to 50 distinct open-ended questions, authored by 79 anonymized students from a single institution across one semester of the academic course 2025-26. It integrates real student answers, instructor-defined ideal solutions, numerical grades and qualitative feedback provided by human teachers, and structured evaluations generated by multiple state-of-the-art LLMs following a unified grading schema. All pedagogical content is available in both Spanish, the original language of instruction, and English, enabling multilingual and cross-lingual research. The dataset supports benchmarking of automated grading systems, analysis of alignment between human and LLM-based assessment, training of judge or meta-evaluation models, and studies on automated feedback generation. By releasing this dataset under an open license, we aim to facilitate transparent, reproducible, and realistic research in educational artificial intelligence and automated assessment. The full description of the characteristics of the dataset, as well as its components, format and methodology, can be found here: "Multilingual Dataset of Student Answers, Human Grading, and Multi-LLM Evaluations for Automated Assessment Research Using JorGPT". Data 2026, 11, 59. https://doi.org/10.3390/data11030059 The dataset is also available on Kaggle: https://www.kaggle.com/datasets/javiersanchezsoriano/jorgpt-student-answers-and-multi-llm-grading

本数据集为一套源自真实大学考核的开放式学生答案多语言数据集。数据集包含79名来自同一院校、修读2025-2026学年某一门课程的匿名学生,在一学期内针对50道不同开放式问题给出的共3041份学生作答内容。数据集整合了真实学生作答、授课教师预设的参考答案、人类教师给出的量化分数与质性反馈,以及多个当前顶尖大语言模型(Large Language Model,简称LLM)按照统一评分框架生成的结构化评估结果。 所有教学相关内容均提供教学原语言西班牙语与英语两个版本,支持多语言及跨语言研究。本数据集可用于自动化评分系统的基准评测、人类评分与大语言模型评分的对齐性分析、评判模型或元评估模型的训练,以及自动化反馈生成相关研究。本数据集以开放许可协议发布,旨在推动教育人工智能与自动化评分领域的透明化、可复现性与贴近真实场景的研究。 该数据集的完整特征说明、组成结构、格式与方法论可参阅论文《"Multilingual Dataset of Student Answers, Human Grading, and Multi-LLM Evaluations for Automated Assessment Research Using JorGPT"》,Data 2026, 11, 59. https://doi.org/10.3390/data11030059。该数据集同时可在Kaggle平台获取:https://www.kaggle.com/datasets/javiersanchezsoriano/jorgpt-student-answers-and-multi-llm-grading

提供机构:
Zenodo
创建时间:
2026-03-12
二维码
社区交流群
二维码
科研交流群
商业服务