Eukalyptus skriven svenska
收藏DataCite Commons2025-12-12 更新2025-04-16 收录
下载链接:
https://spraakbanken.gu.se/resurser/eukalyptus
下载链接
链接失效反馈官方服务:
资源简介:
The Eukalyptus Treebank of Written Swedish is a 100 000 token manually
annotated treebank, consisting of texts from five different genres: novels,
wikipedia, blogs, europarl, and news and community information. The
annotation consists of lemmas, word senses, and parts-of-speech (in part
based on the SUC tagset) and syntactic structures (mainly based on MAMBA and
SAG) which were developed together for the treebank. The Eukalyptus treebank
was developed for evaluation purposes a part of the project Koala - Korps
lingvistiska annotationer, att utveckla en infrastruktur för text-baserad
forskning med högkvalitativa annotationer [Koala – Korp's linguistic
annotations, developing an infrastructure for text-based research with high
quality annotations] funded by Riksbankens Jubileumsfond (2014–2017; nr
In13-0320:1 to Yvonne Adesam et al).
提供机构:
Språkbanken Text
创建时间:
2024-06-18



