This resource contains code to extract a PP attachment disambiguation dataset as described in the paper: Do and Rehbein (2020). Parsers Know Best: German PP Attachment Revisited. The input is in CoNLL
This is a dataset for Indonesian text simplification. Dataset input is provided by Liputan6 dataset (Koto et. al., 2020) and referenced by the id attribute as a document number, and the sentence_index
This video is part of a set of 42 video stimuli designed to study to study variation in the patterns of pronominal marking in the non-Austronesian languages of Alor and Pantar. All of these languages