遇见数据集

SweDia 2000 Research Database

收藏
data.europa2024-06-27 收录
官方服务:

资源简介:

The material was collected in two databases- one for research and one for the general public. The Research Database will be available to researchers in linguistics, Scandinavian languages, etc. The general database is addressed to interested individuals, schools, libraries, etc. Regardless of whether you are a researcher, student, private individual or anything else, the material cannot be used for any other purpose than research and/or education. The research database consists of recordings of a little more than 1300 speakers representing 107 Swedish dialects. The durations of the recordings vary between 25 minutes and over one hour (typically 30 to 40 minutes). Each recording consist of two major parts. One part consisting of controlled material where specific aspects of Swedish phonology are elicited and one part containing spontaneous speech in the form of informal interviews or dialogues between two speakers of the dialect. The controlled material was elicited by the interviewers using crossword like word games in order to avoid any influence of the interviewers own pronunciation of the target words. The material has three parts 1) A wordlist containing words from which the vowel and consonant systems of the dialect can be derived, 2) a wordlist from which the phonological quantity system may be derived and, 3) a list of phrases from which stress and accent patterns may be derived. Typical durations of the controlled and spontaneous material are 10-15 minutes and 15-25 minutes respectively. Database format: The research database is hosted on a server at the Centre for Languages and Literature at Lund University. The meta data standard used for the database is the one defined by the ISLE Meta Data Initiative (IMDI).- utkast Purpose: The purpose is to document and study a large number of Swedish dialects as they sound around the turn of the century 2000. The documentation is done through recordings of dialect speakers for a total of 107 locations in Sweden and in the Swedish-speaking areas of Finland and on Åland. Archiving: The material can be found at the Dialect Archive in Gothenburg (DAG). For more details about DAG, visit: https://www.isof.se/om-oss/kontakt/avdelningen-for-arkiv-och-forskning-i-goteborg.html

本数据集分为两个数据库——分别面向学术研究与普通大众。研究数据库面向语言学、斯堪的纳维亚语研究等领域的科研人员开放;通用数据库则面向有兴趣的个人、学校、图书馆等群体开放。无论您是科研人员、学生、普通个人还是其他身份,本数据集仅可用于学术研究与教育用途,不得挪作他用。 研究数据库包含1300余名发音人的录音数据,涵盖瑞典107种方言。单条录音时长介于25分钟至1小时以上,通常为30至40分钟。每条录音包含两个核心部分:其一为可控采集素材,用于获取瑞典语音学的特定特征信息;其二为自发口语素材,以方言发音人之间的非正式访谈或对话形式呈现。为避免访谈者自身发音对目标词汇的影响,访谈者通过类填字游戏的词汇游戏来采集可控素材。可控素材包含三部分:1)词汇表:收录可用于推导该方言元音与辅音系统的词汇;2)词汇表:收录可用于推导语音音长系统的词汇;3)短语表:收录可用于推导重音与调型模式的短语。可控素材与自发口语素材的典型时长分别为10-15分钟与15-25分钟。 数据库格式:研究数据库部署于隆德大学语言与文学中心的服务器中。该数据库采用的元数据标准由国际口语语料库元数据倡议(ISLE Meta Data Initiative, IMDI)制定。——初稿(utkast) 数据集用途:本数据集旨在记录并研究2000年前后瑞典境内、芬兰瑞典语聚居区以及奥兰群岛的各类瑞典方言语音面貌。本次记录通过对107个采集点的方言发音人进行录音完成,覆盖瑞典全境、芬兰瑞典语聚居区及奥兰群岛。 归档情况:本数据集可在哥德堡方言档案馆(Dialect Archive in Gothenburg, DAG)获取。如需了解该档案馆的更多详情,请访问:https://www.isof.se/om-oss/kontakt/avdelningen-for-arkiv-och-forskning-i-goteborg.html

二维码
社区交流群
二维码
科研交流群
商业服务