Automated detection of software mentions as an instrument for research and infrastructure in the humanities
收藏资源简介:
This dataset contains metadata and external links to full-text publications, along with outputs from Softcite-based automated detection of software mentions. It covers long-running open-access journals in linguistics, literary studies, and digital humanities (DH), selected from DOAJ and organized into two categories: Traditional Linguistics and Literary Studies (TLL) and DH. The DH subset is based on the journal list compiled by Spinaci, Colavizza, and Peroni (https://doi.org/10.5281/zenodo.3406564), and abstracts were retrieved via Crossref. For each record, the dataset provides validated information on the presence of software mentions. It is designed to support diachronic and comparative analyses of software use and methodological change across humanities research communities. The dataset was created within the IBL PAN use case of the SoFAIR project and is associated with the article Automated detection of software mentions as an instrument for research and infrastructure in the humanities. Acknowledgment: This dataset was created as part of the project "Making Software FAIR: A machine-assisted workflow for the research software lifecycle”, under the program CHIST-ERA Open & Re-usable Research Data & Software (CHIST-ERA-22-ORD-08). National Science Center, Poland - Project Number 2022/04/Y/HS2/00182 (Polish Academy of Sciences - IBL-PAN)



