遇见数据集

The CLASSLA-StanfordNLP model for named entity recognition of non-standard Croatian 1.0

收藏
SSH Open MarketPlace2025-07-04 更新2025-07-05 收录
官方服务:

资源简介:

This model for named entity recognition of non-standard Croatian was built with the [CLASSLA-StanfordNLP tool](https://github.com/clarinsi/classla-stanfordnlp) by training on the [hr500k training corpus](http://hdl.handle.net/11356/1183), the [ReLDI-NormTagNER-hr](http://hdl.handle.net/11356/1241) corpus and the [ReLDI-NormTagNER-sr corpus](http://hdl.handle.net/11356/1240), using the [CLARIN.SI-embed.hr word embeddings](http://hdl.handle.net/11356/1205) . The training corpora were additionally augmented for handling missing diacritics by repeating parts of the corpora with diacritics removed. The model is available for download from the CLARIN.SI repository.

创建时间:
2025-07-04
二维码
社区交流群
二维码
科研交流群
商业服务