遇见数据集

Data from 'One question, different annotation depths: A case study in Early Slavic'

收藏
Figshare2022-06-27 更新2026-04-08 收录
官方服务:

资源简介:

This project contains all the datasets and scripts used for the paper: <br> Pedrazzini, Nilo. 2022. One question, different annotation depths: A case study in Early Slavic. <em>Journal of Historical Syntax</em> (Special Collection 'Annotating Historical Corpora') 6(7). 1-40. DOI: 10.18148/hs/2022.v6i4-11.96<br> Content: - das_marianus.csv: all dative absolutes found through TOROT in the Codex Marianus. Used for case study 1 (on 'deeply-annotated treebanks', Section 1 of the paper). - xadvs_marianus.csv: all conjunct participles found through TOROT in the Codex Marianus. Used for case study 1 (on 'deeply-annotated treebanks', Section 1 of the paper). - absdat_nogr.csv: all dative absolutes found through TOROT (except Codex Marianus), as of June 2020. Used for case study 2 (on 'shallowly-annotated treebanks', Section 2 of the paper). - bdinski_da.csv: dative absolutes found in Story of Abraham of Qidun and his niece Mary (Bdinski Sbornik). Used for case study 3 (on 'strategically-annotated treebanks', Section 3 of the paper). - JHS_Pedrazzini.R: R script for all the frequencies and plots in the paper. - harmon.py: script used to harmonize the Church Slavonic and Old East Slavic spellings in case study 2.

提供机构:
Pedrazzini, Nilo
创建时间:
2022-01-07
二维码
社区交流群
二维码
科研交流群
商业服务