Naoero Source Corpus and Cultural-Linguistic Research Dataset
收藏资源简介:
Source-critical, provenance-tracked research dataset accompanying the complete English and Hungarian editions of Naoero: Language, Literature and Cultural Continuity — A Source-Critical Digital Reference. The archive supplies CSV, JSON, JSONL, consolidated JSON and XLSX representations of 16 controlled tables, with 80 sources, 51 prior-art records and 253 bounded provenance edges. It is not a dictionary, normative grammar, community corpus or official orthographic standard, and it has not been validated by native speakers or competent Naoero institutions. This is a MIXED-RIGHTS DATASET. CC BY 4.0 applies only to the project-authored schema, organisation, annotations, classifications, assessments, metadata enrichment and database rights to the extent controlled by Omri Bankuti. Third-party expressions remain under source-specific rights. LICENSE_SCOPE.md controls reuse interpretation; users must also cite the underlying source for reused source material. This is a MIXED-RIGHTS DATASET. The CC BY 4.0 grant applies only to project-authored material to the extent rights are controlled by Omri Bankuti. It does not apply to third-party expressions, source forms, source glosses, third-party translations, scans, recordings, database content, traditional knowledge, Indigenous cultural expressions or other source-specific material. See LICENSE_SCOPE.md and 04_RIGHTS_AND_REUSE.md before reuse. Rights-register rows are non-authoritative due-diligence assessments, not legal determinations. Validation status: SOURCE-CRITICAL, PROVENANCE-TRACKED, NOT NATIVE-VALIDATED.



