遇见数据集

Chinese-Fill-in-the-Blank(CFITB)

收藏
科学数据银行2022-02-15 更新2026-04-23 收录
官方服务:

资源简介:

In order to enrich the Chinese lexical choice data set, taking the people's daily corpus as the initial corpus source, this paper constructs a Chinese test dataset Chinese-Fill-in-the-Blank(CFITB)containing three target word parts of speech: nouns, verbs and adjectives. CFITB dataset contains 500 test samples. Each test sample contains three parts: "number", "test sentence" and "candidate", in which each test sentence contains a target word. Delete the target word and use "__" Instead, the corresponding candidate contains five Chinese words, and each Chinese word is brought into "__" The goal of modeling this dataset is to find the target word corresponding to the most standardized sentence in semantics and grammar from the five candidate sentences.

创建时间:
2021-12-31
二维码
社区交流群
二维码
科研交流群
商业服务