遇见数据集

BenCoref: A Dataset of Nominal Phrases and Pronominal Reference Annotations

收藏
Mendeley Data2021-06-09 更新2026-04-09 收录
官方服务:

资源简介:

This dataset contains 3622 coreference annotations forming 356 coreference clusters in 31630 tokens. The dataset is divided into 69 documents. The text in the documents originate from classic Bengali literature and comes across in 2 categories: short story and novel.

本数据集包含3622条共指标注(coreference annotation),共计在31630个Token(Token)的语料中形成356个共指簇(coreference cluster)。该数据集被划分为69份文档,其文本均源自经典孟加拉语文学作品,且分为短篇小说与长篇小说两类。

创建时间:
2021-06-09
二维码
社区交流群
二维码
科研交流群
商业服务