Paradoxes in the Translation Technology Ecosystem: A Thematically Coded Corpus of Excerpts from the Academic Literature
收藏资源简介:
This dataset contains the coded thematic corpus and analysis outputs for the chapter Paradoxes in the Translation Technology Ecosystem in the Routledge Handbook of Translation and Technology (2nd edition). It comprises 871 verbatim quotes extracted from 59 academic publications using an LLM-assisted coding pipeline with human reliability verification. Quotes are coded against six sensitising paradox codes (C1–C6): Value Distribution, Professional Transformation, Empowerment-Displacement, Automation and Optimisation, Access/Justice/Standardisation, and Data Sovereignty. The deposit includes the consolidated corpus, per-batch coded quote files, analysis output files, methodology documentation, quality logs, and reliability assessment materials used to validate coding quality and analytical consistency. The dataset is intended for research reuse in translation studies, sociotechnical analysis, and AI-in-work scholarship, including secondary analysis of paradox patterns, temporal framing, and code co-occurrence. Access is embargoed until handbook publication. IMPORTANT NOTE ON QUOTE ATTRIBUTIONVerbatim quotes extracted from corpus documents represent the language of the accessed source as it appears on the cited page. In cases where the accessed author is themselves quoting another source, the extracted text may originate with a third party not identified in the dataset metadata. Users citing quotes from this dataset in their own work should verify whether any given passage constitutes the accessed author's own prose or a secondary quotation before attributing it, and should consult the original source document accordingly.



