Environmental large language model Evaluation (ELLE) question-answer (QA) dataset
收藏资源简介:
ELLE-QA数据集是首个专门用于评估大型语言模型在生态和环境科学领域表现的问答数据集。该数据集由北京信息科技大学和清华大学的研究团队创建,包含1130个问答对,涵盖了16个不同的环境主题。数据来源于专家问卷和开源权威材料,确保了数据的专业性和广泛性。数据集通过系统分类,涵盖了内容领域、难度级别和问题类型,旨在为生态和环境领域的AI评估提供一个全面且可靠的框架。该数据集的应用领域包括环境监测、数据分析、教育工具和政策支持,旨在推动生态和环境AI研究的标准化和可持续发展。
The ELLE-QA dataset is the first question-answering dataset specifically designed to evaluate the performance of Large Language Models (LLMs) in the field of ecological and environmental science. Created by research teams from Beijing Information Science and Technology University and Tsinghua University, this dataset contains 1,130 QA pairs covering 16 distinct environmental topics. The data is sourced from expert questionnaires and open-source authoritative materials, ensuring its professionalism and broad coverage. Through systematic classification based on content domains, difficulty levels and question types, the dataset aims to provide a comprehensive and reliable framework for AI evaluation in the ecological and environmental field. Its application areas include environmental monitoring, data analysis, educational tools and policy support, with the goal of promoting the standardization and sustainable development of ecological and environmental AI research.




