遇见数据集

BGRenData -- Bulgarian Text Dataset Manually Annotated with Renarrative for training machine learning models

收藏
Zenodo2025-07-14 更新2026-05-26 收录
官方服务:

资源简介:

This is a dataset created to train classifiers to detect texts in Bulgarian, with forms of Renarrative. When using this dataset, please note that: the errors of the automatic detection of texts with renarrative'' should be taken into account; the presence of a form of "renarrative'' should not be taken as a sole indication of the trustworthiness of the text. When using it, please cite this article: Temnikova, I., Margova, R., Minkov, St., Stefanova, Tsv., Grigorova, N., Gargova, S., Kovatchev, V. (2025) Automatic Detection of the Bulgarian EvidentialRenarrative. Computational Linguistics in Bulgaria. Volume 1.

本数据集专为训练分类器以检测保加利亚语中带有叙事重构(renarrative)形式的文本而构建。 使用本数据集时,请留意以下事项: 需考虑针对带有叙事重构文本的自动检测存在误差; 不应仅以某文本存在叙事重构形式,作为判定该文本可信度的唯一依据。 使用本数据集时,请引用如下文献: Temnikova, I., Margova, R., Minkov, St., Stefanova, Tsv., Grigorova, N., Gargova, S., Kovatchev, V. (2025) 《保加利亚语证据性叙事重构自动检测》。《保加利亚计算语言学》(Computational Linguistics in Bulgaria)第1卷。

提供机构:
Zenodo
创建时间:
2025-07-14
二维码
社区交流群
二维码
科研交流群
商业服务