CO.RA.PAN Metadata (Public)
收藏资源简介:
The CO.RA.PAN Metadata (Public) record provides the openly accessible descriptive layer of the CO.RA.PAN project (Corpus Radiofónico Panhispánico). It offers FAIR-compliant project metadata that enable scientific use, reproducibility, transparent documentation, and bibliographic referencing independently of the restricted corpus data. The public metadata include both global and country-specific information about all recordings in the CO.RA.PAN corpus: program sources, recording dates, durations, technical details, linguistic annotation status, speaker metadata, and corpus-wide variable definitions. These metadata constitute the authoritative description of the corpus composition, structure, annotation, and scope. Contents of this Metadata Repository This record contains the following metadata resources: – corapan_recordings.tsvComplete tabular metadata for all recordings in the corpus. – corapan_recordings.json / corapan_recordings.jsonldMachine-readable global metadata for all recordings. – corapan_corpus_metadata.json / corapan_corpus_metadata.jsonldDefinitions of corpus variables, categories, annotation fields, and structural conventions. – corapan_recordings_{COUNTRYCODE}.json / .jsonldPer-country metadata subsets for individual national corpora. – tei_headers.zipTEI/XML headers describing each transcript segment according to CO.RA.PAN encoding standards. All metadata files are produced and updated through the project’s automated export pipeline and represent the authoritative, versioned documentation of the CO.RA.PAN corpus. No audio data or transcripts are included in this public release.Due to copyright and broadcast restrictions, the full corpus (audio and JSON transcripts) is available only under restricted access in a separate Zenodo dataset. Intended Use This metadata record is designed for: – scientific reference and citation– reproducibility of corpus-based studies– exploration of corpus structure without access to restricted files– integration into external research workflows, catalogues, repositories, and metadata registries The metadata may be reused under the terms of the license associated with this record. FAIR Compliance The CO.RA.PAN metadata follow FAIR principles: – Findable: persistent DOIs for each versioned release– Accessible: openly available TSV, JSON, and JSON-LD files– Interoperable: schema-driven, machine-readable formats compatible with corpus-linguistic tools– Reusable: clear variable definitions, stable naming conventions, permissive CC-BY 4.0 license No audio or transcript data are included here; restrictions applying to the full corpus do not affect FAIR availability of metadata. CO.RA.PAN References and Related Resources CO.RA.PAN Full Corpus (Restricted)DOI: https://doi.org/10.5281/zenodo.15360942 CO.RA.PAN Sample Corpus (Public)DOI: https://doi.org/10.5281/zenodo.15378479 CO.RA.PAN Metadata (Public)DOI: https://doi.org/10.5281/zenodo.17843469 CO.RA.PAN Web ApplicationDOI: https://doi.org/10.5281/zenodo.17834023 Web Application Access CO.RA.PAN Web App (authenticated access to restricted corpus content):https://corapan.online.uni-marburg.de Source code and deployment documentation:https://github.com/FTacke/corapan-webapp Project Overview Further documentation and related digital humanities projects:https://hispanistica.online.uni-marburg.de/



