The Pedagogy Benchmark
收藏资源简介:
该数据集由Fab Inc和AI-for-Education.org创建,旨在评估大型语言模型在跨领域教学知识(CDPK)和特殊教育需求与残疾(SEND)教学知识方面的能力。数据集包含来自智利教育部教师专业发展考试的920个多项选择题,涵盖教学策略和评估方法等教学子领域。这些基准旨在加速和改善大型语言模型在教育应用中的负责任开发,为更有效和基于证据的AI在教育中的应用铺平道路。
This dataset was created by Fab Inc. and AI-for-Education.org, aiming to evaluate the capabilities of large language models (LLMs) in cross-disciplinary pedagogical knowledge (CDPK) and pedagogical knowledge for special educational needs and disabilities (SEND). It includes 920 multiple-choice questions sourced from the Chilean Ministry of Education's teacher professional development examinations, covering pedagogical subfields such as teaching strategies and assessment methods. These benchmarks are designed to accelerate and improve the responsible development of large language models for educational applications, paving the way for more effective and evidence-based AI applications in education.




