BGRenData -- Bulgarian Text Dataset Manually Annotated with Renarrative for training machine learning models
收藏资源简介:
This is a dataset created to train classifiers to detect texts in Bulgarian, with forms of Renarrative. When using this dataset, please note that: the errors of the automatic detection of texts with renarrative'' should be taken into account; the presence of a form of "renarrative'' should not be taken as a sole indication of the trustworthiness of the text. When using it, please cite this article: Temnikova, I., Margova, R., Minkov, St., Stefanova, Tsv., Grigorova, N., Gargova, S., Kovatchev, V. (2025) Automatic Detection of the Bulgarian EvidentialRenarrative. Computational Linguistics in Bulgaria. Volume 1.
本数据集专为训练分类器以检测保加利亚语中带有叙事重构(renarrative)形式的文本而构建。 使用本数据集时,请留意以下事项: 需考虑针对带有叙事重构文本的自动检测存在误差; 不应仅以某文本存在叙事重构形式,作为判定该文本可信度的唯一依据。 使用本数据集时,请引用如下文献: Temnikova, I., Margova, R., Minkov, St., Stefanova, Tsv., Grigorova, N., Gargova, S., Kovatchev, V. (2025) 《保加利亚语证据性叙事重构自动检测》。《保加利亚计算语言学》(Computational Linguistics in Bulgaria)第1卷。



