Paradoxes of Automation in the Translation Technology Ecosystem
收藏资源简介:
This dataset contains the coded thematic corpus and analysis outputs for the chapter Paradoxes in the Translation Technology Ecosystem in the Routledge Handbook of Translation and Technology (2nd edition). It comprises verbatim quotes extracted from academic literature using an LLM-assisted coding pipeline with human reliability verification. Quotes are coded against six sensitising paradox codes (C1–C6): Value Distribution, Professional Transformation, Empowerment-Displacement, Automation and Optimisation, Access/Justice/Standardisation, and Data Sovereignty. The deposit includes the consolidated corpus, per-batch coded quote files, analysis output tables and figures, methodology documentation, quality logs, and reliability assessment materials used to validate coding quality and analytical consistency. The dataset is intended for research reuse in translation studies, sociotechnical analysis, and AI-in-work scholarship, including secondary analysis of paradox patterns, temporal framing, and code co-occurrence. Access is embargoed until handbook publication. IMPORTANT NOTE ON QUOTE ATTRIBUTION Verbatim quotes extracted from corpus documents represent the language of the accessed source as it appears on the cited page. In cases where the accessed author is themselves quoting another source, the extracted text may originate with a third party not identified in the dataset metadata. Users citing quotes from this dataset in their own work should verify whether any given passage constitutes the accessed author's own prose or a secondary quotation before attributing it, and should consult the original source document accordingly.
本数据集收录了《劳特利奇(Routledge)翻译与技术手册(第二版)》中《翻译技术生态系统中的悖论(Paradoxes in the Translation Technology Ecosystem)》一章的编码主题语料库与分析成果。其包含借助大语言模型(LLM)辅助编码流程、经人工信度验证的学术文献逐字引语。 引语将依据六大敏化悖论编码(sensitising paradox codes,C1–C6)进行标注:价值分配、职业转型、赋权-位移、自动化与优化、可及性/公正性/标准化以及数据主权。本数据集存档内容涵盖整合后的语料库、按批次标注的引语文件、分析输出表格与图表、方法学文档、质量日志,以及用于验证编码质量与分析一致性的信度评估材料。 本数据集旨在供翻译研究、社会技术分析以及人工智能与工作相关研究领域复用,包括对悖论模式、时间框架与编码共现情况的二次分析。数据集访问权限将在手册正式出版前处于禁运状态。 关于引语归属的重要说明 从语料库文档中提取的逐字引语,将保留引用页面来源文献的原始语言表述。若所引用的作者本人亦在引用其他来源,则提取文本可能源自数据集元数据中未标注的第三方。使用者若在自身研究中引用本数据集的引语,需先核实指定段落为引用作者的原创文本还是二次引述,并据此查阅原始来源文档。



