BenCoref: A Dataset of Nominal Phrases and Pronominal Reference Annotations
收藏官方服务:
资源简介:
This dataset contains 3622 coreference annotations forming 356 coreference clusters in 31630 tokens. The dataset is divided into 69 documents. The text in the documents originate from classic Bengali literature and comes across in 2 categories: short story and novel.
本数据集包含3622条共指标注(coreference annotation),共计在31630个Token(Token)的语料中形成356个共指簇(coreference cluster)。该数据集被划分为69份文档,其文本均源自经典孟加拉语文学作品,且分为短篇小说与长篇小说两类。
创建时间:
2021-06-09




