CrowdComp Dataset, Course Dataset
收藏资源简介:
CrowdComp数据集包含五个领域:减数分裂、公钥加密、平行公设、牛顿定律和全球变暖。每个CSV文件记录了亚马逊Mechanical Turk上的人类智能任务(HIT)的众包结果。Course数据集包含计算机科学和数学两个领域,文件.edges包含预设概念对,.edges_neg包含该领域的负面示例。
The CrowdComp dataset encompasses five domains: meiosis, public-key cryptography, the parallel postulate, Newton's laws, and global warming. Each CSV file documents the crowdsourced outcomes of human intelligence tasks (HITs) conducted on Amazon Mechanical Turk. The Course dataset includes two fields: computer science and mathematics. The .edges files contain predefined concept pairs, while the .edges_neg files provide negative examples within these domains.
数据集概述
数据集名称
RefD-dataset
数据集来源
用于论文 "Measuring Prerequisite Relations Among Concepts" (Liang et al., 2015) 中的数据。
数据集内容
包含两个子数据集:
-
CrowdComp Dataset
- 包含领域:Meiosis, Public-key Cryptography, Parallel Postulate, Newtons Laws, Global Warming
- 数据格式:每个csv文件记录了Amazon Mechanical Turk上的Human Intelligent Task (HIT)的众包结果。
-
Course Dataset
- 包含领域:Computer Science (CS) 和 Mathematics (MATH)
- 数据格式:.edges 文件包含预设概念对,.edges_neg 文件包含该领域的负例。
- 数据示例:文件中每行格式为 A B,表示B是A的先决条件。例如,CS.edges 中有一行 Network security Computer network,表明 Computer network 是 Network security 的先决条件。
引用信息
如使用此数据,请引用以下论文:
- Liang, Chen et al. "Measuring Prerequisite Relations Among Concepts." Proceedings of the 2015 Conference on Empirical Methods in Natural Language Processing.
- Talukdar, Partha Pratim and Cohen, William W. "Crowdsourced comprehension: predicting prerequisite structure in wikipedia." Proceedings of the Seventh Workshop on Building Educational Applications Using NLP.




