遇见数据集

Datasets to evaluate Generation/Selection of Substitutes tasks to support lexical simplification

收藏
Mendeley Data2021-03-04 更新2026-04-09 收录
官方服务:

资源简介:

These datasets are part of the EASIER corpus and include instances to evaluate the Generation and Selection tasks pertaining to lexical simplification. These datasets consist of a target word and proposed synonyms, along with their necessary metadata, such as the sentence in which the word appears and start and end offsets to locate it. The smallest dataset contains 575 instances in which a word contains three proposed substitutes, while the full dataset contains over 5000 instances with at least one proposed substitute. For more information, address our git repository https://github.com/LURMORENO/EASIER_CORPUS

创建时间:
2021-03-04
二维码
社区交流群
二维码
科研交流群
商业服务