Linguistic Asymmetry Index (LAI): Benchmarking Multilingual Research Infrastructures. Dataset and Reproducible Workflow (v1.0)
收藏资源简介:
This deposit contains the workflow files and derived datasets used to compute the Linguistic Asymmetry Index (LAI) for five digital research infrastructures (CLARIN ERIC, Europeana, OpenAIRE Graph, EUDAT/B2FIND, and Wikidata). The LAI provides a reproducible, metadata-based measure of linguistic asymmetry across five components: language representation, English anchor bias, metadata completeness, institutional concentration, and access inequality. All metadata used in these computations were obtained exclusively from public APIs, OAI-PMH endpoints, or SPARQL services, and no personal or content data were processed. The deposit includes the Python scripts, configuration files, dependency specifications, LAI outputs, and accompanying reports that support the Discussion Paper currently under review at the Journal of Open Humanities Data.



