Semeval-2016 Task 13 Dataset
收藏arXiv2025-09-30 收录
数据链接:
官方服务:
资源简介:
该数据集包含了三个领域的英文数据集,分别是环境、科学和食品。它被用于评估在扩展叶节点方面的性能表现。此外,该数据集在与随机生成的分类法结合使用时,采用了自我监督学习的方式,并对20%的叶节点进行了采样以用于测试。该任务的目标是对叶节点进行分类法的扩展。
This dataset encompasses English datasets across three domains: environmental, scientific, and food. It is employed to evaluate model performance for taxonomy leaf node expansion. Furthermore, when combined with randomly generated taxonomies, this dataset utilizes a self-supervised learning framework, with 20% of the leaf nodes sampled for the test set. The objective of this task is to expand taxonomies by leveraging leaf nodes.
提供机构:
Semeval搜集汇总
数据集介绍

背景与挑战
背景概述
该数据集是SemEval-2016 Task 13(TExEval-2),专注于从文本中自动提取层次关系(如下位词-上位词关系)和构建分类法。它包含多个子任务,如分类法构建和上位词识别,并支持英语、荷兰语、法语和意大利语等多语言环境,旨在评估自然语言处理系统在组织术语结构方面的性能。
以上内容由遇见数据集搜集并总结生成



