遇见数据集

ComPara: A Corpus Linguistics Dataset of Computation in Architecture

收藏
Mendeley Data2020-10-21 更新2026-04-09 收录
官方服务:

资源简介:

A corpus linguistics built to study the language of computational architecture, or architecture which focuses on technology developments. The corpus includes: (1) the volume titles, titles of articles and Introduction keywords for the journal Architectural Design (AD) to retrieve keynotes in theoretical discourse, and (2) titles and abstracts of winning and honourable mentions of the eVolo Skyscraper competition to retrieve words in conceptual project titles and their descriptions. This dataset has around 100.000 words and can serve as a basis for quantitative, qualitative or mixed method analysis of the language used in AD and the eVolo skyscraper competition between 2005 and 2019. As AD is recognized as one of the journals focusing on the 'digital turns' in architecture, and eVolo is arguably the most prestigious architectural competitions which focus on technological advances in architecture, ComPara can be considered respresentative for the language of computational architecture over the last 15 years. It includes .txt and .csv files as well as .svg wordclouds.

本语料库专为语料库语言学研究构建,旨在探究计算建筑学(computational architecture,即聚焦技术发展的建筑学分支)的语言使用特征。该语料库包含两部分内容:(1) 《建筑设计》(Architectural Design,简称AD)期刊的卷名、文章标题及引言关键词,用于提取理论论述中的核心主旨;(2) eVolo摩天大楼竞赛获奖与荣誉提名作品的标题及摘要,用于获取概念性项目标题及其描述文本中的词汇。本数据集总词汇量约10万,可作为2005年至2019年间《建筑设计》与eVolo摩天大楼竞赛相关文本语言的定量、定性或混合方法分析的基础。鉴于《建筑设计》是公认的聚焦建筑学“数字转向”的核心期刊之一,而eVolo竞赛无疑是当前最具声望的聚焦建筑学技术革新的建筑赛事,本语料库ComPara可视为近15年来计算建筑学领域语言使用情况的代表性样本。该数据集包含纯文本(.txt)、逗号分隔值(.csv)文件以及可缩放矢量图形(.svg)格式的词云图。

创建时间:
2020-10-21
二维码
社区交流群
二维码
科研交流群
商业服务