遇见数据集

Metadata for a Corpus of English as a Lingua Franca Medical Discourse in Latvia (2019–2025)

收藏
Zenodo2026-02-24 更新2026-05-26 收录
官方服务:

资源简介:

Corpus Description and Documentation This repository contains metadata for a specialised corpus of professional medical discourse in English as a Lingua Franca (ELF), compiled for the study of medical specialists’ monologic interactional competence in Latvia. The corpus consists of 200,099 tokens of transcribed spoken discourse produced by Latvian medical professionals in international and professional communication settings between 2019 and 2025. The data represent three institutional genres: scientific conference presentations, public panel discussions, and professional interviews. All interactions involve extended monologic or semi-monologic speech produced for professional audiences. All recordings were transcribed with the help of the professional transcription tool Notta, followed by manual verification. Basic transcription conventions were used, focusing on lexical content and discourse organisation rather than detailed phonetic representation. Anonymisation and ethical considerations To protect participants’ privacy and prevent potential identification within a relatively small professional community: speakers were anonymised and assigned coded identifiers, event titles and contextual details were generalised or rephrased, institutional affiliations and personal names were removed. Due to ethical and confidentiality considerations, full transcripts and audio recordings are not publicly available. The repository provides detailed metadata describing corpus composition. Anonymised excerpts and additional documentation may be made available upon reasonable request. The study was conducted in accordance with the Code of Ethics for Scientists of Latvia and the Academic Ethics Codex of the University of Latvia.

提供机构:
Zenodo
创建时间:
2026-02-24
二维码
社区交流群
二维码
科研交流群
商业服务