NER4Legal SRB
收藏资源简介:
NER4Legal SRB数据集是由塞尔维亚人工智能研究所创建,包含75份塞尔维亚上诉法院的裁决书,这些裁决书经过仔细挑选,以提供塞尔维亚司法实践的代表样本。数据集以官方 gazette 的公告、法律、法院裁决等8种命名实体类型进行了字符级别的标注,采用BIO标注方案,形成了总共2172个包含命名实体的句子。该数据集可用于塞尔维亚法律文书中命名实体的识别研究。
The NER4Legal SRB dataset was created by the Artificial Intelligence Research Institute of Serbia. It consists of 75 rulings from Serbian appellate courts, which were carefully selected to provide a representative sample of Serbian judicial practice. The dataset is annotated at the character level with 8 types of named entities including official gazette announcements, laws, court rulings and others using the BIO annotation scheme, resulting in a total of 2172 sentences containing named entities. This dataset can be used for research on named entity recognition in Serbian legal documents.




