Textmining Von Konferenzabstracts: Dokumentation Eines Arbeitsprozesses In Den Digitalen Geisteswissenschaften. Ein Werkstattbericht
收藏资源简介:
This dataset accompanies a research of the mining of conference abstracts. The article describes the digital workflow of automatically extracting various entities from a corpus of conference abstracts and performing network analysis on the results in order to find relationships. The aim is to identify the spread and diversity of tools, the representation of institutions and federations, and the thematic and personal networks of speakers at a German-speaking DH conference. Particular emphasis is placed on the documentation of the entire process, which can be interpreted as an initiative to develop interdisciplinary standards in this field. Included in the dataset are a training corpus (Trainingskorpuserweitert.csv) a configuration file (Dariah6.prop) a model file gained from the Stanford parser (dariah—6ner-model2.ser.gz)



